Lab health
Is the lab healthy?
The machine behind every number on this site: is it awake, is it still signed in, and what did the last round of checks do?
Running means the last round started under 180 minutes ago. A few rate limits a day is normal on paid plans.
If the lab is behind, read the boards as stale and check back. Reading this site never starts a test — the schedule is ours alone.
Next check
—A new round starts every hour, day and night. The next one is at 4:00 AM.
How fresh is this?
When the lab last ran, and how far behind it is right now.
- Last speed round
4m ago · 13 saved · 0 skipped · 2 failed · took 4m 46s
The hourly ping and speed sweep
- Most recent problem
overloaded
The newest error label in the last 24 hours
- What counts on the boards
Test kit v0.2 only · the answer must pass our grader · leftover plugins, skills, or hooks are named on the sample
The rule the live boards apply to every sample
Where each plan's meter stands
How full our own usage window is on each paid plan right now, when it resets, and how solid today's reading is. This is our account, not yours.
- Claude Max 20x
- 71% of the window usedGrade — · Gaps
Resets Thu 04:00 UTC · 16 readings that day · read 3h ago
- ChatGPT Pro ($200, 20×)
- 4.0% of the window usedGrade — · Gaps
Resets Mon 00:41 UTC · 1687 readings that day · read 3h ago
- SuperGrok Heavy
- 37% of the window usedGrade — · Gaps
Resets Sun 17:53 UTC · 12 readings that day · read 3h ago
We read our own account's meter from each vendor's usage endpoint, a few times an hour. It costs no model calls, and it says nothing about your account.
24-hour tape
The last 24 hours
Every ping, lined up round by round. This is the app waking up — not a writing-speed test.
Claude Code
Anthropic
3.21s
Start-to-finish this round
100%
Pings that worked in 24h
31% fasterthan its usual 24-hour time
Start-to-finish · 24 hours
- Usual
- 4.69s
- Slowest
- 8.36s
- Finished
- 24/24
Codex
OpenAI
5.94s
Start-to-finish this round
100%
Pings that worked in 24h
7% fasterthan its usual 24-hour time
Start-to-finish · 24 hours
- Usual
- 6.40s
- Slowest
- 18.7s
- Finished
- 21/21
Grok Build
xAI
4.63s
Start-to-finish this round
100%
Pings that worked in 24h
12% fasterthan its usual 24-hour time
Start-to-finish · 24 hours
- Usual
- 5.28s
- Slowest
- 17.7s
- Finished
- 24/24
How to read this
Ping is a tiny ok check. Wall time is start to finish — lower is better. In the chart, a red hashed stub failed, and a bar that fades at the top ran past the scale — hover it for the real number. Writing speed lives on the home leaderboard, in visible answer tok/s.
- Rounds with a check
- 24 of 24
- Each round is hour
- Checks saved
- 69
- 0 failed
- Checks that worked
- 100.0%
- Finished with no error
Wall time · last 24 hours
One bar per agent, every hour
Every column is one slot on the same clock, so the bars line up across agents.
How to read this
Each one-hour slot has one bar per agent, side by side. Height is how long the ping took, so a shorter bar is faster. A red hatched stub is a probe that failed. A bar that fades at the top ran past the scale — hover it for the real number. A thin flat line means no probe ran that slot.
Full record — every slot, every number
| Hour | Claude Code | Codex | Grok Build | |||
|---|---|---|---|---|---|---|
| Wall time | Tok/s | Wall time | Tok/s | Wall time | Tok/s | |
| Aug 24 3:00am | 3.21s | 12.45 | 5.94s | 0.84 | 4.63s | 3.67 |
| Aug 24 2:00am | 3.45s | 22.34 | 5.78s | 0.87 | 5.21s | 3.26 |
| Aug 24 1:00am | 4.75s | 13.89 | 18.7s | 0.27 | 6.90s | 4.35 |
| Aug 24 12:00am | 4.69s | 25.18 | 7.06s | 0.71 | 6.15s | 4.88 |
| Aug 23 11:00pm | 4.14s | 34.27 | 6.20s | 0.81 | 4.46s | 3.81 |
| Aug 23 10:00pm | 5.63s | 6.93 | 18.5s | 0.27 | 6.39s | 4.70 |
| Aug 23 9:00pm | 4.58s | 25.09 | 12.0s | 0.42 | 4.22s | 7.11 |
| Aug 23 8:00pm | 4.83s | 15.93 | 11.3s | 0.44 | 17.7s | 0.96 |
| Aug 23 7:00pm | 4.50s | 18.89 | 9.60s | 0.52 | 4.70s | 6.18 |
| Aug 23 6:00pm | 3.19s | 20.70 | no sample | 8.01s | 3.62 | |
| Aug 23 5:00pm | 3.48s | 21.57 | no sample | 6.61s | 4.39 | |
| Aug 23 4:00pm | 4.37s | 17.84 | no sample | 7.28s | 3.98 | |
| Aug 23 3:00pm | 4.94s | 14.77 | 8.52s | 0.59 | 5.80s | 2.93 |
| Aug 23 2:00pm | 5.14s | 20.64 | 6.40s | 0.78 | 5.19s | 5.59 |
| Aug 23 1:00pm | 4.69s | 22.17 | 6.54s | 0.76 | 5.23s | 3.25 |
| Aug 23 12:00pm | 4.73s | 13.12 | 10.1s | 0.50 | 4.79s | 3.76 |
| Aug 23 11:00am | 4.90s | 19.61 | 4.73s | 1.06 | 15.3s | 1.90 |
| Aug 23 10:00am | 7.62s | 16.41 | 4.72s | 1.06 | 5.26s | 5.52 |
| Aug 23 9:00am | 8.36s | 7.06 | 4.76s | 1.05 | 5.47s | 5.49 |
| Aug 23 8:00am | 3.96s | 16.67 | 5.30s | 0.94 | 7.23s | 2.49 |
| Aug 23 7:00am | 4.95s | 16.16 | 12.1s | 0.41 | 5.31s | 3.20 |
| Aug 23 6:00am | 5.72s | 12.23 | 5.69s | 0.88 | 5.11s | 5.68 |
| Aug 23 5:00am | 4.37s | 20.82 | 4.97s | 1.01 | 4.58s | 3.71 |
| Aug 23 4:00am | 4.39s | 10.71 | 4.63s | 1.08 | 4.62s | 3.68 |
Last check 4m ago
69 records in view · all times America/New_York
The apps we are signed into
We buy the same paid plans you can buy. If a sign-in goes stale, we skip that app for the round and say so here, rather than publishing a broken number.
- Claude Code sign-in
Signed in · claude.ai
claude.ai · max · v2.1.241 (Claude Code)
- Codex sign-in
Signed in · ChatGPT
chatgpt · vcodex-cli 0.147.0
- Grok Build sign-in
Signed in · grok.com
session · vgrok 1.0.5 (5115b46bc9) [stable]
- No API keys
None set — we test the paid apps only
An API key in the environment would quietly move billing to the pay-as-you-go API, and we would stop measuring the paid app people actually use. Our runner strips those keys, and stops on purpose if one is set.
Test kit v0.2 · win32. App versions change often — a jump can move the numbers.
What the last round did
Every model the last speed round touched, in order. Recorded means we saved a measurement. Skipped means it was not that model’s turn, or the app had already hit a rate limit.
Show every target from the last round (13)
- ping · codex
- saved · ok · ok · 5.94s
- ping · grok
- saved · ok · ok · 4.63s
- ping · claude
- saved · ok · ok · 3.21s
- speed · claude/fable
- saved · failed · rate limit warning · 7.82s
- speed · codex/luna
- saved · ok · ok · 9.87s
- speed · codex/terra
- saved · ok · ok · 9.54s
- speed · codex/sol
- saved · ok · ok · 10.8s
- speed · codex/gpt-5.5
- saved · ok · ok · 10.3s
- speed · grok/grok-4.6
- saved · ok · ok · 5.95s
- speed · grok/grok-4.5
- saved · ok · ok · 6.53s
- speed · claude/haiku
- saved · ok · ok · 6.00s
- speed · claude/sonnet
- saved · ok · ok · 4.50s
- speed · claude/opus
- saved · failed · overloaded · 3m 15s
Read next
- Logs
Every sample and every fail behind these counters.
- How we test
The rules this machine follows on every run.
- Quota
What a full week of each plan is worth.
- Live board
The numbers this lab is producing right now.