Lab health
Is the lab healthy?
The machine behind every number on this site: is it awake, is it still signed in, and what did the last round of checks do?
Running means the last round started under 180 minutes ago. A few rate limits a day is normal on paid plans.
If the lab is behind, read the boards as stale and check back. Reading this site never starts a test — the schedule is ours alone.
Next check
—A new round starts every hour, day and night. The next one is at 6:00 PM.
How fresh is this?
When the lab last ran, and how far behind it is right now.
- Last speed round
1m ago · 13 saved · 0 skipped · 6 failed · took 1m 25s
The hourly ping and speed sweep
- Most recent problem
failed
The newest error label in the last 24 hours
- What counts on the boards
Test kit v0.2 only · the answer must pass our grader · leftover plugins, skills, or hooks are named on the sample
The rule the live boards apply to every sample
Where each plan's meter stands
How full our own usage window is on each paid plan right now, when it resets, and how solid today's reading is. This is our account, not yours.
- Claude Max 20x
- 59% of the window usedGrade C · Thin
Resets Thu 04:00 UTC · 326 readings that day · read 17h ago
- ChatGPT Pro ($100, 5×)
- 49% of the window usedGrade — · Gaps
Resets Tue 00:30 UTC · 7 readings that day · read 17h ago
- SuperGrok Heavy
- 17% of the window usedGrade — · Gaps
Resets Sun 17:53 UTC · 14 readings that day · read 17h ago
We read our own account's meter from each vendor's usage endpoint, a few times an hour. It costs no model calls, and it says nothing about your account.
24-hour tape
The last 24 hours
Every ping, lined up round by round. This is the app waking up — not a writing-speed test.
Claude Code
Anthropic
3.48s
Start-to-finish this round
100%
Pings that worked in 24h
23% fasterthan its usual 24-hour time
Start-to-finish · 24 hours
- Usual
- 4.54s
- Slowest
- 8.36s
- Finished
- 24/24
Codex
OpenAI
—
Start-to-finish this round
100%
Pings that worked in 24h
Waiting on this round’s check
Start-to-finish · 24 hours
- Usual
- 6.24s
- Slowest
- 12.1s
- Finished
- 22/22
Grok Build
xAI
6.61s
Start-to-finish this round
100%
Pings that worked in 24h
27% slowerthan its usual 24-hour time
Start-to-finish · 24 hours
- Usual
- 5.20s
- Slowest
- 15.3s
- Finished
- 24/24
How to read this
Ping is a tiny ok check. Wall time is start to finish — lower is better. In the chart, a red hashed stub failed, and a bar that fades at the top ran past the scale — hover it for the real number. Writing speed lives on the home leaderboard, in visible answer tok/s.
- Rounds with a check
- 24 of 24
- Each round is hour
- Checks saved
- 70
- 0 failed
- Checks that worked
- 100.0%
- Finished with no error
Wall time · last 24 hours
One bar per agent, every hour
Every column is one slot on the same clock, so the bars line up across agents.
How to read this
Each one-hour slot has one bar per agent, side by side. Height is how long the ping took, so a shorter bar is faster. A red hatched stub is a probe that failed. A bar that fades at the top ran past the scale — hover it for the real number. A thin flat line means no probe ran that slot.
Full record — every slot, every number
| Hour | Claude Code | Codex | Grok Build | |||
|---|---|---|---|---|---|---|
| Wall time | Tok/s | Wall time | Tok/s | Wall time | Tok/s | |
| Aug 23 5:00pm | 3.48s | 21.57 | no sample | 6.61s | 4.39 | |
| Aug 23 4:00pm | 4.37s | 17.84 | no sample | 7.28s | 3.98 | |
| Aug 23 3:00pm | 4.94s | 14.77 | 8.52s | 0.59 | 5.80s | 2.93 |
| Aug 23 2:00pm | 5.14s | 20.64 | 6.40s | 0.78 | 5.19s | 5.59 |
| Aug 23 1:00pm | 4.69s | 22.17 | 6.54s | 0.76 | 5.23s | 3.25 |
| Aug 23 12:00pm | 4.73s | 13.12 | 10.1s | 0.50 | 4.79s | 3.76 |
| Aug 23 11:00am | 4.90s | 19.61 | 4.73s | 1.06 | 15.3s | 1.90 |
| Aug 23 10:00am | 7.62s | 16.41 | 4.72s | 1.06 | 5.26s | 5.52 |
| Aug 23 9:00am | 8.36s | 7.06 | 4.76s | 1.05 | 5.47s | 5.49 |
| Aug 23 8:00am | 3.96s | 16.67 | 5.30s | 0.94 | 7.23s | 2.49 |
| Aug 23 7:00am | 4.95s | 16.16 | 12.1s | 0.41 | 5.31s | 3.20 |
| Aug 23 6:00am | 5.72s | 12.23 | 5.69s | 0.88 | 5.11s | 5.68 |
| Aug 23 5:00am | 4.37s | 20.82 | 4.97s | 1.01 | 4.58s | 3.71 |
| Aug 23 4:00am | 4.39s | 10.71 | 4.63s | 1.08 | 4.62s | 3.68 |
| Aug 23 3:00am | 4.10s | 20.73 | 6.08s | 0.82 | 6.94s | 4.18 |
| Aug 23 2:00am | 5.58s | 16.14 | 4.64s | 1.08 | 5.21s | 3.26 |
| Aug 23 1:00am | 7.77s | 13.13 | 5.34s | 0.94 | 4.64s | 6.47 |
| Aug 23 12:00am | 3.76s | 11.71 | 6.67s | 0.75 | 4.57s | 3.72 |
| Aug 22 11:00pm | 3.60s | 12.21 | 6.53s | 0.77 | 4.43s | 3.84 |
| Aug 22 10:00pm | 3.10s | 29.37 | 5.93s | 0.84 | 4.96s | 5.85 |
| Aug 22 9:00pm | 4.11s | 31.17 | 7.32s | 0.68 | 5.17s | 3.29 |
| Aug 22 8:00pm | 2.87s | 28.55 | 7.34s | 0.68 | 5.21s | 3.26 |
| Aug 22 7:00pm | 2.63s | 15.21 | 8.58s | 0.58 | 5.04s | 5.96 |
| Aug 22 6:00pm | 5.44s | 24.46 | 8.15s | 0.61 | 4.85s | 5.98 |
Last check 1m ago
70 records in view · all times America/New_York
The apps we are signed into
We buy the same paid plans you can buy. If a sign-in goes stale, we skip that app for the round and say so here, rather than publishing a broken number.
- Claude Code sign-in
Signed in · claude.ai
claude.ai · max · v2.1.241 (Claude Code)
- Codex sign-in
Signed in · ChatGPT
chatgpt · vcodex-cli 0.147.0
- Grok Build sign-in
Signed in · grok.com
session · vgrok 1.0.5 (5115b46bc9) [stable]
- No API keys
None set — we test the paid apps only
An API key in the environment would quietly move billing to the pay-as-you-go API, and we would stop measuring the paid app people actually use. Our runner strips those keys, and stops on purpose if one is set.
Test kit v0.2 · win32. App versions change often — a jump can move the numbers.
What the last round did
Every model the last speed round touched, in order. Recorded means we saved a measurement. Skipped means it was not that model’s turn, or the app had already hit a rate limit.
Show every target from the last round (13)
- ping · claude
- saved · ok · ok · 3.48s
- ping · codex
- saved · failed · failed · 6.91s
- ping · grok
- saved · ok · ok · 6.61s
- speed · claude/fable
- saved · failed · rate limit warning · 6.46s
- speed · codex/luna
- saved · failed · failed · 7.09s
- speed · codex/terra
- saved · failed · failed · 6.42s
- speed · codex/sol
- saved · failed · failed · 5.96s
- speed · codex/gpt-5.5
- saved · failed · failed · 6.33s
- speed · grok/grok-4.6
- saved · ok · ok · 7.25s
- speed · grok/grok-4.5
- saved · ok · ok · 6.25s
- speed · claude/haiku
- saved · ok · ok · 6.62s
- speed · claude/sonnet
- saved · ok · ok · 5.24s
- speed · claude/opus
- saved · ok · ok · 4.80s
Read next
- Logs
Every sample and every fail behind these counters.
- How we test
The rules this machine follows on every run.
- Quota
What a full week of each plan is worth.
- Live board
The numbers this lab is producing right now.