Skip to content
BetaBenchAlert v0.1 is in Beta.Numbers are real, but pages and rules can still change.See what changed

The rest of the field

Other models

The 6 pins that run every round but do not lead the home page. Same prompts, same hourly pulse, same units as the 4 models on the main boards.

The models we lead with

Fable 5, Opus 5, GPT-5.6 Sol, and Grok 4.6 lead the main boards on the home page.

Open the live board

Speed

Typical speed

These 6 models, ranked by their typical total wait in the last 24 hours. Writing speed is shown beside it.

Writing speed counts only the answer you can see. Total wait measures click-to-finish time. The table ranks the shorter total wait first.

The 6 models outside the home boards, ranked by typical start-to-finish time over the last 24 hours.
RankModelBar, on one shared scaleWriting speedTotal waitvs yesterdayAnsweredCoverage
1Grok 4.5xAI · effort high38.176.26slikely 5.88s6.54s-9%100%24/24
2Sonnet 5Anthropic · effort high34.517.00slikely 6.23s7.33s-3%100%24/24
3Haiku 4.5Anthropic · effort high28.288.67slikely 6.75s9.31s+1%100%24/24
4GPT-5.6 TerraOpenAI · effort high · last round failed23.7710.2slikely 9.64s11.0s-3%83%24/24
5GPT-5.5OpenAI · effort high · last round failed23.9210.2slikely 9.81s10.8s+8%92%24/24
6GPT-5.6 LunaOpenAI · effort high · last round failed23.0410.6slikely 9.79s12.4s-6%92%24/24

Window 24 hours · 6 of 6 models have 12 completed checks · 144 checks attempted · test kit v0.2. A dash under vs yesterday means that model did not have 12 completed checks the day before. 6 of 10 pinned models shown · the 4 we lead with are on the live board.

1 retired pin

Retired pins

Pins we can no longer run. Their pages stay up, and their old rounds stay in Logs.