Skip to content
BetaBenchAlert v0.1 is in Beta.Numbers are real, but pages and rules can still change.See what changed

Last 24 hours

Today’s report

Who finished fastest, what a week of each plan is worth, and the ranked board.

BenchAlert

Aug 25, 2026

Sonnet 5

4.82stypical total wait

50.2 tok/s

Opus 5 completed the check 2.1× faster than GPT-5.6 Sol.

Featured

  1. 1Opus 55.14s46.9 tok/s+34%
  2. 2Grok 4.66.37s37.5 tok/s
  3. 3GPT-5.6 Sol11.0s22.1 tok/s
  4. Fable 5 tok/s

benchalert.com

last 24 hours · typical total wait + answer tok/s

Quota

What 100% of a weekly limit is worth at the vendor's own API prices.

How we measure
QuotaAPI $ / week

Liveupdated 6h agonext run 03:10 UTC

  1. Claude Max 20xB≈ $2,342

    5-hour ≈ $1,013 · Fable 5 up to 50% of weekly

  2. ChatGPT Pro ($200, 20×)degraded≈ $2,313
  3. SuperGrok Heavydegraded≈ $1,627

    Includes Cursor Ultra — $400/mo usage on top

Last reading
Aug 25
Method
v0.2
Interval
daily

Speed

Typical speed and total wait

All 10 pinned models, ranked by typical total wait over the last 24 hours. At least 12 completed checks are required.

Writing speed counts only the answer you can see. Total wait measures click-to-finish time. The table ranks the shorter total wait first.

All 10 pinned models, ranked by typical start-to-finish time over the last 24 hours, with visible writing speed beside it.
RankModelBar, on one shared scaleWriting speedTotal waitvs yesterdayAnsweredCoverage
1Sonnet 5Anthropic · effort high50.224.82slikely 4.58s5.26s+18%96%23/24
2Opus 5Anthropic · effort high46.895.14slikely 4.70s5.91s+34%100%24/24
3Grok 4.5xAI · effort high39.436.06slikely 5.52s6.89s+5%100%24/24
4Grok 4.6xAI · effort xhigh37.536.37slikely 5.99s6.77s+1%100%24/24
5Haiku 4.5Anthropic · effort high38.076.42slikely 6.22s6.91s+26%92%22/24
6GPT-5.6 TerraOpenAI · effort high23.2410.5slikely 9.73s11.3s0%100%25/24
7GPT-5.6 LunaOpenAI · effort high22.4310.9slikely 9.92s12.1s+8%100%24/24
8GPT-5.6 SolOpenAI · effort high22.1011.0slikely 10.00s12.6s-1%100%24/24
9GPT-5.5OpenAI · effort high21.6711.3slikely 10.3s11.9s0%100%24/24
Fable 5Anthropic · effort high · collecting 10/12 · last round rate limit42%16/24

Window 24 hours · 9 of 10 models have 12 completed checks · 230 checks attempted · test kit v0.2. A dash under vs yesterday means that model did not have 12 completed checks the day before. 4 of 10 pinned models shown.

24 hours

The last 24 hours, round by round

Each line joins one model's rounds, one an hour. Every dot is a real round, so the swings stay in plain sight. Point at the plot to read any round.

Visible answer tokens per second · higher is faster

Model writing speed over time

Writing speed (tokens / second)020406080
8am2pm8pm2amnow
Aug 24Aug 25

Lab time (ET)

Line: one model's visible writing speed, check to check. Dots: the checks themselves. A single missed hour is stepped over by a faint dotted link; anything longer breaks the line. Nothing is invented to fill a gap.

Full record — every round, every number
Visible answer tokens per second for every check in the last 24 hours. A dash means that check had no completed answer.
RoundOpus 5Grok 4.6GPT-5.6 SolFable 5
Aug 25 6:00am52.0636.5122.15
Aug 25 5:00am49.0840.8824.30
Aug 25 4:00am34.3241.1525.43
Aug 25 3:00am48.0936.7317.67
Aug 25 2:00am30.9025.2616.09
Aug 25 1:00am37.4935.7024.71
Aug 25 12:00am42.6937.5823.35
Aug 24 11:00pm50.9332.0121.48
Aug 24 10:00pm40.7632.6718.18
Aug 24 9:00pm56.4136.1020.54
Aug 24 8:00pm51.2138.0124.94
Aug 24 7:00pm51.8741.7022.08
Aug 24 6:00pm51.9242.9222.12
Aug 24 5:00pm42.4139.8924.99
Aug 24 4:00pm39.2640.3424.6723.89
Aug 24 3:00pm45.6839.8818.3826.67
Aug 24 2:00pm39.8238.7721.3047.79
Aug 24 1:00pm48.9839.6519.2547.25
Aug 24 12:00pm45.3935.2724.3240.79
Aug 24 11:00am35.7121.0618.4343.60
Aug 24 10:00am40.8341.5122.9925.15
Aug 24 9:00am51.2435.3223.7428.07
Aug 24 8:00am54.3337.4821.6542.78
Aug 24 7:00am54.4431.3315.7042.83

Window 24 hours · one round every 60 minutes · 24 rounds on the clock · n = 82 completed readings drawn · test kit v0.2. 4 of 10 pinned models shown.

We also ping each app every round to check it answers at all. What that check asks lives on How we test.

How to read this

Writing speed counts only the visible answer over start-to-finish time. Everything here comes from one lab machine on test kit v0.2 — not an official vendor number.