Claude Sonnet 5.5vsClaude Fable 5.1
SI Score, benchmarks, price and context compared, with every number sourced.
79.5
SI Score
#5 of 124 ranked models
79.4
SI Score
#6 of 124 ranked models
Pillar by pillar
Claude Sonnet 5.5Claude Fable 5.1
86.4 Reasoning25% of score 78.9
90.6 Math25% of score 92.3
61.6 Coding25% of score 63.9
79.5 Preference25% of score 82.5
The basics
| Attribute | Claude Sonnet 5.5 | Claude Fable 5.1 |
|---|---|---|
| Input price, per 1M tokens | $2.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | $10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ |
| Output price, per 1M tokens | $10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | $50.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ |
| Context window | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ |
| Released | Sep 28, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | Sep 1, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ |
| Weights | Closed | Closed |
Highlighted values are the lower price or the larger context window.
Shared benchmarks
22 in common. Best result ahead: Claude Sonnet 5.5 on 6, Claude Fable 5.1 on 13Each side shows its best published result. Results can come from different settings or harnesses, so open a row to compare like with like before reading much into a small gap.
FrontierMath Tier 4 (v2) 80.5% 87.8%
Claude Sonnet 5.5
- max effort80.5%Epoch AI BenchmarkingEpoch-owned evaluation mean score; scores only, no benchmark questions [variant] maxPublished Sep 29, 2026
Retrieved Oct 9, 2026 · CC-BY
Open source ↗
Claude Fable 5.1
- max effort87.8%Epoch AI BenchmarkingEpoch-owned evaluation mean score; scores only, no benchmark questions [variant] maxPublished Sep 1, 2026
Retrieved Oct 9, 2026 · CC-BY
Open source ↗
FrontierMath Tiers 1–3 (v2) 88.8% 90.2%
Claude Sonnet 5.5
- max effort88.8%Epoch AI BenchmarkingEpoch-owned evaluation mean score; scores only, no benchmark questions [variant] maxPublished Sep 29, 2026
Retrieved Oct 9, 2026 · CC-BY
Open source ↗
Claude Fable 5.1
- max effort90.2%Epoch AI BenchmarkingEpoch-owned evaluation mean score; scores only, no benchmark questions [variant] maxPublished Sep 1, 2026
Retrieved Oct 9, 2026 · CC-BY
Open source ↗
Humanity's Last Exam (with tools) 64.5% 65%
Claude Sonnet 5.5
- Published result64.5%Official model cards via models.devLab-reported; metric score; transcribed by MIT models.dev catalog; not independently evaluated [variant] with toolsPublished Sep 28, 2026
Retrieved Oct 9, 2026 · factual citation; MIT transcription
Open source ↗
Claude Fable 5.1
- Published result65%Official model cards via models.devLab-reported; metric score; transcribed by MIT models.dev catalog; not independently evaluated [variant] with tools; production safeguards with fallbackPublished Sep 1, 2026
Retrieved Oct 9, 2026 · factual citation; MIT transcription
Open source ↗
LiveBench Coding: code completion 84.8% 82.6%
Claude Sonnet 5.5
- Published result84.8%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result82.6%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Coding: code generation 93.0% 90.1%
Claude Sonnet 5.5
- Published result93.0%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result90.1%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Coding: JavaScript 36.4% 68.2%
Claude Sonnet 5.5
- Published result36.4%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result68.2%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Coding: Python 35% 70%
Claude Sonnet 5.5
- Published result35%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result70%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Coding: TypeScript 46.7% 60%
Claude Sonnet 5.5
- Published result46.7%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result60%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Math: AMPS Hard 98% 99%
Claude Sonnet 5.5
- Published result98%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result99%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Math: competition math 97.1% 97.1%
Claude Sonnet 5.5
- Published result97.1%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result97.1%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Math: integrals 100% 99%
Claude Sonnet 5.5
- Published result100%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result99%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Math: olympiad 91.7% 93.0%
Claude Sonnet 5.5
- Published result91.7%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result93.0%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Math: simplify 72.4% 70.2%
Claude Sonnet 5.5
- Published result72.4%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result70.2%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Reasoning: connections 98.5% 99.3%
Claude Sonnet 5.5
- Published result98.5%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result99.3%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Reasoning: consecutive events 90.9% 90.6%
Claude Sonnet 5.5
- Published result90.9%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result90.6%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Reasoning: logic with navigation 80% 86%
Claude Sonnet 5.5
- Published result80%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result86%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Reasoning: spatial 98% 100%
Claude Sonnet 5.5
- Published result98%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result100%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Reasoning: theory of mind 69.2% 80.8%
Claude Sonnet 5.5
- Published result69.2%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result80.8%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LiveBench Reasoning: zebra puzzles 100% 100%
Claude Sonnet 5.5
- Published result100%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
Claude Fable 5.1
- Published result100%LiveBenchLiveBench subtask result; dataset edition 2026_06_25; publication time is HTTP Last-Modified of the score CSV, not the dataset edition or individual evaluation date [variant] 2026_06_25Published Oct 7, 2026
Retrieved Oct 9, 2026 · factual citation; Apache-2.0 code
Open source ↗
LMArena Text 1471.2 elo 1510.0 elo
Claude Sonnet 5.5
- Published result1471.2 eloLMArena / ArenaPublished source fact [variant] text / overallPublished Oct 8, 2026
Retrieved Oct 9, 2026 · CC-BY-4.0
Open source ↗
Claude Fable 5.1
- Published result1510.0 eloLMArena / ArenaPublished source fact [variant] text / overallPublished Oct 8, 2026
Retrieved Oct 9, 2026 · CC-BY-4.0
Open source ↗
OTIS Mock AIME 2024–2025 100% 100%
Claude Sonnet 5.5
- max effort100%Epoch AI BenchmarkingEpoch-owned evaluation mean score; scores only, no benchmark questions [variant] maxPublished Sep 29, 2026
Retrieved Oct 9, 2026 · CC-BY
Open source ↗
Claude Fable 5.1
- max effort100%Epoch AI BenchmarkingEpoch-owned evaluation mean score; scores only, no benchmark questions [variant] maxPublished Sep 1, 2026
Retrieved Oct 9, 2026 · CC-BY
Open source ↗
Terminal-Bench 4.0 70.6% 57.9%
Claude Sonnet 5.5
- Setting 170.6%Official model cards via models.devLab-reported; metric score; transcribed by MIT models.dev catalog; not independently evaluated [variant] 4.0Published Sep 28, 2026
Retrieved Oct 9, 2026 · factual citation; MIT transcription
Open source ↗ - Setting 261.8%Terminal-BenchTerminal-Bench 4.0; published harness submission, 95% CI retained at source [variant] Claude Code; maxPublished Oct 6, 2026
Retrieved Oct 9, 2026 · Apache-2.0; factual citation
Open source ↗
Claude Fable 5.1
- Setting 157.9%Terminal-BenchTerminal-Bench 4.0; published harness submission, 95% CI retained at source [variant] Claude Code; maxPublished Sep 3, 2026
Retrieved Oct 9, 2026 · Apache-2.0; factual citation
Open source ↗ - Setting 257.9%Terminal-BenchTerminal-Bench 4.0; published harness submission, 95% CI retained at source [variant] Claude Code; xhighPublished Sep 17, 2026
Retrieved Oct 9, 2026 · Apache-2.0; factual citation
Open source ↗ - Setting 355.8%Official model cards via models.devLab-reported; metric score; transcribed by MIT models.dev catalog; not independently evaluated [variant] production safeguards with fallback; 4.0Published Sep 1, 2026
Retrieved Oct 9, 2026 · factual citation; MIT transcription
Open source ↗ - Setting 454.5%Terminal-BenchTerminal-Bench 4.0; published harness submission, 95% CI retained at source [variant] Claude Code; highPublished Sep 17, 2026
Retrieved Oct 9, 2026 · Apache-2.0; factual citation
Open source ↗ - Setting 553.9%Terminal-BenchTerminal-Bench 4.0; published harness submission, 95% CI retained at source [variant] Claude Code; mediumPublished Sep 17, 2026
Retrieved Oct 9, 2026 · Apache-2.0; factual citation
Open source ↗ - Setting 643.3%Terminal-BenchTerminal-Bench 4.0; published harness submission, 95% CI retained at source [variant] Claude Code; lowPublished Sep 17, 2026
Retrieved Oct 9, 2026 · Apache-2.0; factual citation
Open source ↗
Which should you choose?
- For reasoning, Claude Sonnet 5.5 leads by 7.5 points.
- Claude Sonnet 5.5 costs less per input token ($2.00 vs $10.00 per 1M).
- Confidence is 93% for Claude Sonnet 5.5 and 100% for Claude Fable 5.1; sources still to report can move either score.
These follow from the numbers above. They're not a verdict on your use case.