Updated Oct 9, 2026. 142 models ranked, 58 benchmarks, 29 sources
The superintelligence leaderboard
Every frontier AI model ranked on one score built only from public benchmarks, with every number traced to its source.
- Leads overall Claude Fable 5.1 80.2SI Score
- Best at coding DeepSeek V4.1 Flash 73.3coding pillar
- Best at math GPT-6.1 Sol 93.0math pillar
- Best at reasoning Claude Opus 5.5 91.4reasoning pillar
- Best open weights Kimi K3 71.0SI Score
- Cheapest in the top 20 Gemini 3.7 Flash $1.50per 1M, blended
- Longest context Llama 4 Scout 17B Instruct 10Mtokens
Top 10 AI models
Ranked by SI Score, 0 to 100, from 58 public benchmarks. Pillar cells glow brighter with higher scores. How the score works
| # | Model | SI Score | Coding | Math | Reasoning | Preference | Priceper 1M, blended | Context | Released |
|---|---|---|---|---|---|---|---|---|---|
| 1 | Claude Fable 5.1 Anthropic | 80.2 | Coding 70 | Math 91 | Reasoning 87 | Preference 82 | $20.00 | 1M | Sep 1, 2026 |
| 2 | Claude Opus 5.5 Anthropic | 78.4 | Coding 72 | Math 92 | Reasoning 91 | Preference 83 | $8.00 | 1M | Sep 22, 2026 |
| 3 | GPT-6 Astra OpenAI | 77.5 | Coding 65 | Math 93 | Reasoning 87 | Preference 77 | $20.00 | 1.1M | Sep 3, 2026 |
| 4 | Claude Fable 5 Anthropic | 76.8 | Coding 70 | Math 91 | Reasoning 89 | Preference 81 | $20.00 | 1M | Jun 9, 2026 |
| 5 | Claude Opus 5 Anthropic | 74.9 | Coding 70 | Math 87 | Reasoning 84 | Preference 82 | $10.00 | 1M | Jul 24, 2026 |
| 6 | GPT-6.1 Sol OpenAI | 73.0 | Coding 65 | Math 93 | Reasoning 88 | Preference 77 | $4.00 | 1.1M | Sep 29, 2026 |
| 7 | GPT-5.6 Sol OpenAI | 72.1 | Coding 64 | Math 89 | Reasoning 82 | Preference 78 | $8.00 | 1.1M | Jul 9, 2026 |
| 8 | Claude Opus 4.7 Anthropic | 71.2 | Coding 66 | Math 78 | Reasoning 74 | Preference 80 | $10.00 | 1M | Apr 16, 2026 |
| 9 | Kimi K3 Moonshot AIOpen weights | 71.0 | Coding 71 | Math 74 | Reasoning 81 | Preference 80 | $6.00 | 1M | Jul 16, 2026 |
| 10 | Claude Opus 4.6 Anthropic | 70.8 | Coding 62 | Math 74 | Reasoning 79 | Preference 82 | $10.00 | 1M | Feb 5, 2026 |
Intelligence index
SI Score of the top 25 models, colored by lab. Hatched bars have open weights.
Best by task
Pillar leaders, among models with at least three results. Price is blended per 1M tokens.
Intelligence vs price
Blended API price per 1M tokens, log scale. Up and left is more score per dollar.
Labs
Each lab's best ranked model.
Latest releases
Newest text models in the catalog.
- Oct 7 Claude Haiku 5.5Anthropic 59.4#64
- Oct 6 Mistral Large 4Mistral AI 62.2#53
- Sep 29 Ling 3.1 FlashinclusionAI Awaiting results
- Sep 29 GPT-6.1 SolOpenAI 73.0#6
- Sep 28 Claude Sonnet 5.5Anthropic 69.4#15
- Sep 27 MiniMax-M3.1-Flash-PreviewMiniMax Awaiting results
- Sep 25 LongCat-2.5-PreviewMeituan Awaiting results
Price, top 10
Blended price per 1M tokens (3 input : 1 output).
Best value
SI Score per blended dollar.
Context window
Largest advertised input, log scale.
The frontier
Best SI Score among models released by each date.
AI news
Model launches from the labs and moves on our leaderboard.
- Oct 7 Introducing Claude Haiku 5.5 on AWS Amazon AWS Machine LearningClaude Haiku 5.5 · #64 · SI 59.4
- Oct 7 Claude Haiku 5.5 now ranks #64 SuperIndex leaderboardClaude Haiku 5.5 · #64 · SI 59.4
- Oct 6 Introducing Mistral Large 4 Mistral AIMistral Large 4 · #53 · SI 62.2
- Oct 6 Mistral Large 4 now ranks #53 SuperIndex leaderboardMistral Large 4 · #53 · SI 62.2
- Sep 29 Introducing GPT-6.1 Sol OpenAIGPT-6.1 Sol · #6 · SI 73.0
- Sep 29 Ling 3.1 Flash was released (not yet ranked) SuperIndex leaderboard
- Sep 29 GPT-6.1 Sol now ranks #6 SuperIndex leaderboardGPT-6.1 Sol · #6 · SI 73.0
Open weights vs proprietary
The best downloadable model trails the best closed model by 9.2 points.
Benchmarks
Most-reported benchmarks and who leads each.
- LMArena Textpreference · 125 models Claude Opus 5.582.8
- GPQA Diamondreasoning · 119 models GPT-6 Astra95.8
- OTIS Mock AIME 2024–2025math · 110 models Claude Fable 5.1100.0
- FrontierMath Tiers 1–3 (v2)math · 70 models GPT-6 Astra93.7
- ARC-AGI-2 (semi-private)reasoning · 66 models Claude Opus 5.593.3
- ARC-AGI-1 (semi-private)reasoning · 65 models Claude Opus 5.598.5
Hear when the frontier moves
A notification on this device when a new model takes #1 or enters the top 10. No account, no email; turn it off any time.
Understand the numbers
Methodology
How the SI Score is built, what the confidence % means, and which sources feed it, with their licenses.
What is superintelligence?
The research term, the corporate lane name and the new government spelling, untangled and dated.
“Super Intelligence” in U.S. policy
Executive Order 14434 made it executive-branch vocabulary on Sep 29, 2026. What it says, verbatim.
Frontier models API
The current frontier per lab as free JSON, for agents and developers. No key, no signup.