Weekly LLM benchmark digest
GPT-5.6 arrives, locked behind a preview
OpenAI previewed GPT-5.6 Sol, Terra, and Luna on June 26, undercutting Anthropic’s flagship by half on price. The catch: it shipped to roughly 20 government-approved partners under a limited preview, so most teams could not use it yet. Claude Mythos 5 still led at 99, with Fable 5 at 95, while Z.AI’s GLM-5.2 entered the top five at 91. The issue examined what the 17-day gap between those launches suggested about the market.
Leaderboard movers
Claude Mythos 5
Anthropic · 99 overall
#1
Held #1
Claude Fable 5
Anthropic · 95 overall
#2
Held #2
GLM-5.2
Z.AI · 91 overall
#4
New to top five
Analysis worth opening
This issue in numbers
- 124
- Models ranked
- 3
- New that week
- $23.92
- Average output price per 1M tokens
- 99
- Top overall score