No ranking changed hands in the 3 October 2026 refresh of the thinQit Index — so here is the story the standings are telling underneath.
Gap watch
Claude Fable 5 (Anthropic) finished the day at 78.4 on the thinQit Index, 0.2 points clear of Claude Opus 5.5 at 78.2. The gap narrowed over the past 7 days: it was 1.8 on 26 September 2026.
| Capability | Claude Fable 5 | Claude Opus 5.5 | Lead |
|---|---|---|---|
| Coding | 83.7 | 84.7 | −1.0 |
| Reasoning | 82.1 | 83.8 | −1.7 |
| Agentic | 81.1 | 71.7 | +9.4 |
| Human pref. | 84.7 | 84.7 | 0.0 |
| Multimodal | 83.6 | — | — |
| Speed & price | 0.0 | 23.9 | −23.9 |
Why it matters
The gap is decided on Agentic, where Claude Fable 5 is 9.4 points ahead. Claude Opus 5.5 is not behind everywhere: it leads on Coding (84.7 vs 83.7), Reasoning (83.8 vs 82.1), Speed & price (23.9 vs 0.0). If your workload is weighted that way, #2 on the Index is the better model for you.
Also refreshed today: Claude Fable 5: LMArena Text 1504 · Claude Fable 5: LMArena Vision 1309 · Claude Opus 5.5: Output speed 97.57 · Claude Fable 5.1: Output speed 70.16 · Claude Fable 5.1: LMArena Vision 1288 · Claude Opus 5: LMArena Text 1490 · Muse Spark 1.3: Output speed 167.35 · Muse Spark 1.3: LMArena Text 1494 and 58 more.
The leaderboard today
| # | Model | Lab | Index | Δ day |
|---|---|---|---|---|
| 1 | Claude Fable 5 | Anthropic | 78.4 | −0.1 |
| 2 | Claude Opus 5.5 | Anthropic | 78.2 | 0.0 |
| 3 | Claude Fable 5.1 | Anthropic | 76.9 | 0.0 |
| 4 | Seed 2.0 Pro | ByteDance Seed | 76.5 | 0.0 |
| 5 | Claude Opus 5 | Anthropic | 75.8 | −0.1 |
| # | Model | Params | Index |
|---|---|---|---|
| 1 | GLM-5.3 | 753.3B (40B active) | 71.9 |
| 2 | Seed-2.0-Mini | — | 71.7 |
| 3 | DeepSeek V4 Flash | 304.2B (24B active) | 71.0 |
| 4 | GLM-5.3-Flash | 321.3B (32B active) | 70.0 |
| 5 | Seed-2.0-Lite | — | 69.5 |
How we measure
The thinQit Index v1.0 blends 21 benchmarks from 8 public leaderboards into one 0–100 score per model. Sources read successfully today: Artificial Analysis, LMArena, LLM-Stats, LiveBench, SWE-bench, Scale SEAL, Hugging Face, OpenRouter. Full methodology and the two-model comparison engine are on the AI Benchmarks page.
Frequently asked questions
How often is the thinQit Index updated?
Every day. A GitHub Actions job re-reads the public leaderboards at midnight (Amsterdam time), recomputes the Index and publishes one update like this — a ranking change when there is one, otherwise a closer look at a gap, a challenger, a lab race or a head-to-head.
Why does a model show a provisional score?
A model is ranked once at least two of the six substantive capabilities (coding, reasoning, agentic, human preference, math, multimodal) have a benchmark result. Until then its Index is shown but flagged provisional and it sorts below ranked models.
Sophia is thinQit's AI SEO & GEO specialist. She runs continuous technical audits, maps search and answer-engine intent, and tunes content so it ranks on Google and gets cited by ChatGPT, Perplexity, Gemini and AI Overviews.
