Updated Oct 9, 2026

The superintelligence leaderboard

Every frontier AI model ranked on one score, built only from public benchmarks. Refreshed daily, with every number traced to its source.

See the rankings

The leaderboard

Full catalog

142 text models ranked from 58 benchmarks across 29 sources. Scores run 0 to 100; confidence shows how much of the expected evidence has reported. How the SI Score works

1 Claude Fable 5.1 Anthropic 80.2 100% confidence 100 percent, Full $10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$50.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 1, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
2 Claude Opus 5.5 Anthropic 78.4 100% confidence 100 percent, Full $4.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$20.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 22, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
3 GPT-6 Astra OpenAI 77.5 100% confidence 100 percent, Full $10.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
$50.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 3, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
4 Claude Fable 5 Anthropic 76.8 100% confidence 100 percent, Full $10.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$50.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
5 Claude Opus 5 Anthropic 74.9 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 24, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
6 GPT-6.1 Sol OpenAI 73.0 100% confidence 100 percent, Full $2.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
$10.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 29, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
7 GPT-5.6 Sol OpenAI 72.1 100% confidence 100 percent, Full $4.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
$20.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
8 Claude Opus 4.7 Anthropic 71.2 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 16, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
9 Kimi K3 Moonshot AI 71.0 100% confidence 100 percent, Full $3.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k3 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 16, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
10 Claude Opus 4.6 Anthropic 70.8 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 5, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
11 Muse Spark 1.3 Meta 70.1 93% confidence 93 percent, High $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
12 Gemini 3.7 Flash Google 70.0 100% confidence 100 percent, Full $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 13, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
13 GPT-5.4 OpenAI 69.7 100% confidence 100 percent, Full $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4 Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4 Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 5, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
14 GPT-5.5 OpenAI 69.5 100% confidence 100 percent, Full $5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
15 Claude Sonnet 5.5 Anthropic 69.4 93% confidence 93 percent, High $2.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 28, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
16 GPT-6 Sol OpenAI 68.9 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
17 Gemini 3.8 Flash Google 68.8 100% confidence 100 percent, Full $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 2, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
18 Qwen3.8 Max Alibaba / Qwen 68.7 85% confidence 85 percent, High $1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.95Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 3, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
19 Gemini 3.5 Flash Google 68.3 100% confidence 100 percent, Full $1.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$9.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 19, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
20 Claude Sonnet 4.6 Anthropic 68.3 100% confidence 100 percent, Full $3.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 17, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
21 GPT-5.6 Terra OpenAI 68.2 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-terra Retrieved Oct 9, 2026 · MIT
Open source ↗
$12.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-terra Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
22 Gemini 3.1 Pro Preview Google 68.1 100% confidence 100 percent, Full $2.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$12.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 19, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
23 MiniMax-M3 MiniMax 68.0 100% confidence 100 percent, Full $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 1, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
24 GLM-5.3 Z.ai 67.7 93% confidence 93 percent, High $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 14, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
25 Claude Sonnet 5 Anthropic 67.7 93% confidence 93 percent, High $2.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$10.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 30, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
26 Claude Opus 4.8 Anthropic 67.5 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 28, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
27 Grok 4.6 xAI 66.6 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
$6.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
500Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
28 Inkling Thinking Machines Lab 66.5 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
29 Qwen3.6 Max Preview Alibaba / Qwen 66.4 64% confidence 64 percent, Medium $1.30Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤128K; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$7.80Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤128K; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 20, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
30 GLM-5.3-Flash Z.ai 66.2 100% confidence 100 percent, Full $0.15Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.50Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 26, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
31 DeepSeek V4.1 Flash DeepSeek 66.2 69% confidence 69 percent, Medium $0.15DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.60DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 10, 2026DeepSeek V4.1 Flash announcementOfficial dated introduction and availability announcement Retrieved Oct 9, 2026 · factual citation
Open source ↗
Open
32 Kimi K2 Thinking Turbo Moonshot AI 66.1 64% confidence 64 percent, Medium not yet reported not yet reported 262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 6, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
33 GPT-5.2 OpenAI 65.9 100% confidence 100 percent, Full $1.75models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.2 Retrieved Oct 9, 2026 · MIT
Open source ↗
$14.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.2 Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 11, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
34 GLM-5.2 Z.ai 65.7 100% confidence 100 percent, Full $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 13, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
35 Gemini 3.6 Flash Google 65.5 100% confidence 100 percent, Full $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 21, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
36 Grok 4.7 xAI 65.5 100% confidence 100 percent, Full $2.00xAI models & pricingPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
$6.00xAI models & pricingPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
500KxAI models & pricingPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 21, 2026xAI models & pricingOfficial featured model page datePublished; model introduction date Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
37 Grok 4.5 xAI 64.8 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$6.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
500Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 8, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
38 Kimi K2.6 Moonshot AI 64.0 85% confidence 85 percent, High $0.95models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k2.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k2.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 21, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
39 Qwen3.5 397B-A17B Alibaba / Qwen 64.0 69% confidence 69 percent, Medium $0.17Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.03Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
40 GPT-5.5 Pro OpenAI 63.5 53% confidence 53 percent, Medium $30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$180.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
41 GPT-5.6 Luna OpenAI 63.4 100% confidence 100 percent, Full $0.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-luna Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-luna Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
42 Gemma 4 26B A4B IT Google 63.4 64% confidence 64 percent, Medium $0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
43 GLM-5.1 Z.ai 63.3 64% confidence 64 percent, Medium $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 7, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
44 DeepSeek V4 Pro DeepSeek 63.1 85% confidence 85 percent, High $1.32LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$3.96LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
45 Gemma 4 31B IT Google 63.0 64% confidence 64 percent, Medium $0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
46 DeepSeek V4 Pro 0813 DeepSeek 63.0 69% confidence 69 percent, Medium $0.66DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.98DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
47 Muse Spark 1.1 Meta 62.9 61% confidence 61 percent, Medium $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 8, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
48 Qwen3.7 Plus Alibaba / Qwen 62.7 64% confidence 64 percent, Medium $0.28Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.10Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
49 Qwen3.5 35B-A3B Alibaba / Qwen 62.7 64% confidence 64 percent, Medium $0.06Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.46Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 23, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
50 Claude Opus 4.5 Anthropic 62.6 100% confidence 100 percent, Full $5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-opus-4-5-20251101 Retrieved Oct 9, 2026 · MIT
Open source ↗
$25.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-opus-4-5-20251101 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
51 MiMo-V2.5-Pro Xiaomi 62.6 88% confidence 88 percent, High $0.43models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.87models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
52 Muse Spark 1.2 Meta 62.3 53% confidence 53 percent, Medium $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
53 Mistral Large 4 Mistral AI 62.2 64% confidence 64 percent, Medium $0.68Mistral API pricingOfficial Mistral Serverless API Standard rate, displayed sale price where applicable; cache, Batch, specialist units and hosted third-party models excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$2.09Mistral API pricingOfficial Mistral Serverless API Standard rate, displayed sale price where applicable; cache, Batch, specialist units and hosted third-party models excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 6, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
54 Qwen3.8 27B Alibaba / Qwen 61.9 69% confidence 69 percent, Medium $0.50Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$3.00Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 14, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
55 GPT-6 Luna OpenAI 61.6 100% confidence 100 percent, Full $0.10OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.50OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
56 Qwen3.6 Plus Alibaba / Qwen 61.6 80% confidence 80 percent, High $0.28Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
57 GLM-4.7 Z.ai 61.3 69% confidence 69 percent, Medium $0.60Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$2.20Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 22, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
58 GPT-5.4 Pro OpenAI 61.3 69% confidence 69 percent, Medium $30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$180.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 5, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
59 Nemotron 3 Ultra 550B A55B NVIDIA 61.2 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 4, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
60 DeepSeek V4 Flash 0731 DeepSeek 61.0 69% confidence 69 percent, Medium not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 31, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
61 Hy3 Tencent 60.7 88% confidence 88 percent, High $0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://cloud.tencent.com/document/product/1823/130050. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; tencent-tokenhub/hy3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://cloud.tencent.com/document/product/1823/130050. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; tencent-tokenhub/hy3 Retrieved Oct 9, 2026 · MIT
Open source ↗
256Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 6, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
62 Claude Haiku 4.5 Anthropic 60.5 64% confidence 64 percent, Medium $1.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-haiku-4-5-20251001 Retrieved Oct 9, 2026 · MIT
Open source ↗
$5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-haiku-4-5-20251001 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 15, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
63 Qwen3.5 Flash Alibaba / Qwen 59.5 64% confidence 64 percent, Medium $0.03Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.29Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 23, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
64 Claude Haiku 5.5 Anthropic 59.4 68% confidence 68 percent, Medium $0.10Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.50Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Oct 7, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
65 GLM-5 Z.ai 59.4 90% confidence 90 percent, High $1.00Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$3.20Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
66 Kimi K2.5 Moonshot AI 59.0 100% confidence 100 percent, Full $0.60LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://platform.moonshot.ai/docs/guide/kimi-k2-5-quickstart. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$3.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://platform.moonshot.ai/docs/guide/kimi-k2-5-quickstart. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 1, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
67 Qwen3.7 Max Alibaba / Qwen 58.6 53% confidence 53 percent, Medium $1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.95Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 21, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
68 Gemini 3 Pro Preview Google 58.6 68% confidence 68 percent, Medium not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 18, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
69 GPT-5.5 Instant OpenAI 58.3 64% confidence 64 percent, Medium not yet reported not yet reported 400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 5, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
70 MiniMax-M2.7 MiniMax 57.9 88% confidence 88 percent, High $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 18, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
71 Fugu Ultra Sakana AI 57.9 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
72 MiniMax-M2.5 MiniMax 57.8 100% confidence 100 percent, Full $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
73 Grok 4.20 (Reasoning) xAI 57.7 80% confidence 80 percent, High $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.20-0309-reasoning Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.20-0309-reasoning Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
74 MiMo-V2.6-Pro Xiaomi 57.4 88% confidence 88 percent, High $0.43models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.87models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
75 GPT-5.4 nano OpenAI 57.4 100% confidence 100 percent, Full $0.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 17, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
76 GPT-5.4 mini OpenAI 56.9 100% confidence 100 percent, Full $0.75models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 17, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
77 LongCat-2.0 Meituan 56.6 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 30, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
78 MiMo-V2.5 Xiaomi 56.4 88% confidence 88 percent, High $0.14models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.28models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
79 DeepSeek V4 Flash DeepSeek 56.2 53% confidence 53 percent, Medium $0.30LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.20LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
80 Fugu Sakana AI 55.9 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
81 MiMo-V2.6-Flash Xiaomi 55.8 88% confidence 88 percent, High $0.14models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.28models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
82 DeepSeek V3 0324 DeepSeek 55.7 74% confidence 74 percent, Medium not yet reported not yet reported 164Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 24, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
83 GPT-5 OpenAI 55.7 100% confidence 100 percent, Full $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5 Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 7, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
84 MiniMax-M2 MiniMax 55.7 96% confidence 96 percent, High $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 27, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
85 GPT-5.1 OpenAI 55.7 85% confidence 85 percent, High $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 13, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
86 Qwen3.6 27B Alibaba / Qwen 55.6 53% confidence 53 percent, Medium $0.60Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$3.60Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
87 Claude Opus 4 Anthropic 55.5 100% confidence 100 percent, Full $15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$75.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 22, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
88 o3 OpenAI 55.4 87% confidence 87 percent, High $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/o3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$8.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/o3 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 16, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
89 Gemini 3.5 Flash Lite Google 55.2 100% confidence 100 percent, Full $0.30Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$2.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 21, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
90 GPT OSS 120B OpenAI 54.8 79% confidence 79 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
91 Gemini 3 Flash Preview Google 54.8 93% confidence 93 percent, High $0.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 17, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
92 Grok 4.3 xAI 54.6 80% confidence 80 percent, High $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.3 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 17, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
93 DeepSeek-V3 DeepSeek 54.1 93% confidence 93 percent, High $0.27LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 26, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
94 Claude Sonnet 4.5 Anthropic 53.9 80% confidence 80 percent, High $3.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-sonnet-4-5-20250929 Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-sonnet-4-5-20250929 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 29, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
95 Inkling Small Thinking Machines Lab 53.6 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 30, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
96 DeepSeek V3.2 DeepSeek 53.3 56% confidence 56 percent, Medium $0.28LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
97 GLM-4.7-Flash Z.ai 53.3 69% confidence 69 percent, Medium $0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.z.ai/guides/overview/pricing. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; zai/glm-4.7-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.z.ai/guides/overview/pricing. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; zai/glm-4.7-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 19, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
98 Step 3.5 Flash StepFun 53.0 88% confidence 88 percent, High $0.10models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.stepfun.com/docs/zh/overview/concept. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; stepfun/step-3.5-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.30models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.stepfun.com/docs/zh/overview/concept. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; stepfun/step-3.5-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
256Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 29, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
99 Qwen3 30B A3B Alibaba / Qwen 52.6 64% confidence 64 percent, Medium $0.11Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.43Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 28, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
100 GPT OSS 20B OpenAI 52.5 64% confidence 64 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
101 Claude Sonnet 3.5 v2 Anthropic 51.9 100% confidence 100 percent, Full not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 22, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
102 Claude Opus 4.1 Anthropic 51.5 80% confidence 80 percent, High not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
103 o4-mini OpenAI 51.2 87% confidence 87 percent, High $1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 16, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
104 Llama 4 Scout 17B Instruct Meta 51.0 64% confidence 64 percent, Medium not yet reported not yet reported 10Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
105 GPT-5 Mini OpenAI 50.6 99% confidence 99 percent, High $0.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 7, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
106 Qwen3 32B Alibaba / Qwen 50.4 67% confidence 67 percent, Medium $0.16Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.64Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
107 Qwen3 235B-A22B Alibaba / Qwen 50.3 76% confidence 76 percent, Medium $0.29Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.15Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
108 Qwen3 Max Alibaba / Qwen 50.2 69% confidence 69 percent, Medium $0.36Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤32K; Non-Thinking mode only. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.43Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤32K; Non-Thinking mode only. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 23, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
109 GPT-4o (2024-05-13) OpenAI 49.4 93% confidence 93 percent, High $5.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 13, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
110 Claude Sonnet 3.7 Anthropic 48.9 100% confidence 100 percent, Full not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 19, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
111 Mistral Small 3.1 24B Mistral AI 48.4 64% confidence 64 percent, Medium not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 17, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
112 QwQ 32B Alibaba / Qwen 48.2 67% confidence 67 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
113 DeepSeek-R1 DeepSeek 47.9 88% confidence 88 percent, High $0.55LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.19LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 20, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
114 GLM-4.6 Z.ai 47.8 55% confidence 55 percent, Medium $0.60Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$2.20Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 30, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
115 Claude Sonnet 4 Anthropic 46.6 100% confidence 100 percent, Full $3.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 22, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
116 Nemotron 3.5 Lightning 30B A3B NVIDIA 46.1 100% confidence 100 percent, Full not yet reported not yet reported 262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 11, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
117 GPT-5 Pro OpenAI 45.8 64% confidence 64 percent, Medium $15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$120.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 6, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
118 Llama-3.1-70B-Instruct Meta 45.7 93% confidence 93 percent, High not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 23, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
119 Gemini 2.5 Pro Google 45.2 90% confidence 90 percent, High $1.25Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$10.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 5, 2025Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
120 Gemma 3 12B IT Google 45.1 64% confidence 64 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 12, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
121 GPT-4o (2024-11-20) OpenAI 42.2 61% confidence 61 percent, Medium $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-11-20 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-11-20 Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 20, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
122 Gemma 3 27B IT Google 42.0 74% confidence 74 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 12, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
123 o3-mini OpenAI 41.6 96% confidence 96 percent, High $1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 20, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
124 GPT-4.1 OpenAI 41.5 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
$8.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 14, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
125 Claude Haiku 3.5 Anthropic 40.8 100% confidence 100 percent, Full $0.80Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 22, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
126 Claude Haiku 3 Anthropic 39.9 93% confidence 93 percent, High not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 13, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
127 Qwen2.5-Coder-32B-Instruct Alibaba / Qwen 39.6 90% confidence 90 percent, High not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 12, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
128 Gemma 3 4B IT Google 39.2 64% confidence 64 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 12, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
129 GPT-5 Nano OpenAI 39.1 85% confidence 85 percent, High $0.05models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.40models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 7, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
130 Llama-3.1-8B-Instruct Meta 37.7 93% confidence 93 percent, High not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 23, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
131 Nova Pro Amazon 37.4 100% confidence 100 percent, Full not yet reported not yet reported 300Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 3, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
132 GPT-4.1 mini OpenAI 36.9 87% confidence 87 percent, High $0.40models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.60models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 14, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
133 GPT-4o (2024-08-06) OpenAI 36.7 100% confidence 100 percent, Full $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-08-06 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-08-06 Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 6, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
134 Nova Lite Amazon 36.6 100% confidence 100 percent, Full not yet reported not yet reported 300Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 3, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
135 Mistral Large 2.1 Mistral AI 35.3 100% confidence 100 percent, Full not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 18, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
136 Mistral Medium 3 Mistral AI 33.0 85% confidence 85 percent, High not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 7, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
137 Llama-3.3-70B-Instruct Meta 32.7 100% confidence 100 percent, Full not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 6, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
138 Llama-3.2-1B Meta 32.1 93% confidence 93 percent, High not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 25, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
139 GPT-4.1 nano OpenAI 32.1 82% confidence 82 percent, High $0.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 14, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
140 GPT-4o OpenAI 31.2 59% confidence 59 percent, Medium $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 13, 2024OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
141 Llama 4 Maverick 17B Instruct Meta 30.0 78% confidence 78 percent, Medium not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
142 GPT-4o mini OpenAI 22.9 100% confidence 100 percent, Full $0.15models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.60models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 18, 2024OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—

Ranked text models appear first. The provisional toggle includes only text models awaiting evidence. Browse every kind in the model catalog. Prices are USD per 1M tokens (official first-party API). Dotted values carry their source: hover or tap to see it. “Not yet reported” means the source hasn’t published that figure for this model.

The frontier

Release timeline

Best SI Score among models released by each date. Each dot is a ranked model at its current score.

20 40 60 80 20252026 Claude Haiku 3: SI 39.9, released Mar 2024GPT-4o (2024-05-13): SI 49.4, released May 2024GPT-4o: SI 31.2, released May 2024GPT-4o mini: SI 22.9, released Jul 2024Llama-3.1-70B-Instruct: SI 45.7, released Jul 2024Llama-3.1-8B-Instruct: SI 37.7, released Jul 2024GPT-4o (2024-08-06): SI 36.7, released Aug 2024Llama-3.2-1B: SI 32.1, released Sep 2024Claude Sonnet 3.5 v2: SI 51.9, released Oct 2024Claude Haiku 3.5: SI 40.8, released Oct 2024Qwen2.5-Coder-32B-Instruct: SI 39.6, released Nov 2024Mistral Large 2.1: SI 35.3, released Nov 2024GPT-4o (2024-11-20): SI 42.2, released Nov 2024Nova Pro: SI 37.4, released Dec 2024Nova Lite: SI 36.6, released Dec 2024Llama-3.3-70B-Instruct: SI 32.7, released Dec 2024o3-mini: SI 41.6, released Dec 2024DeepSeek-V3: SI 54.1, released Dec 2024DeepSeek-R1: SI 47.9, released Jan 2025Claude Sonnet 3.7: SI 48.9, released Feb 2025QwQ 32B: SI 48.2, released Mar 2025Gemma 3 12B IT: SI 45.1, released Mar 2025Gemma 3 27B IT: SI 42.0, released Mar 2025Gemma 3 4B IT: SI 39.2, released Mar 2025Mistral Small 3.1 24B: SI 48.4, released Mar 2025DeepSeek V3 0324: SI 55.7, released Mar 2025Qwen3 32B: SI 50.4, released Apr 2025Qwen3 235B-A22B: SI 50.3, released Apr 2025Llama 4 Scout 17B Instruct: SI 51.0, released Apr 2025Llama 4 Maverick 17B Instruct: SI 30.0, released Apr 2025GPT-4.1: SI 41.5, released Apr 2025GPT-4.1 mini: SI 36.9, released Apr 2025GPT-4.1 nano: SI 32.1, released Apr 2025o3: SI 55.4, released Apr 2025o4-mini: SI 51.2, released Apr 2025Qwen3 30B A3B: SI 52.6, released Apr 2025Mistral Medium 3: SI 33.0, released May 2025Claude Opus 4: SI 55.5, released May 2025Claude Sonnet 4: SI 46.6, released May 2025Gemini 2.5 Pro: SI 45.2, released Jun 2025GPT OSS 120B: SI 54.8, released Aug 2025GPT OSS 20B: SI 52.5, released Aug 2025Claude Opus 4.1: SI 51.5, released Aug 2025GPT-5: SI 55.7, released Aug 2025GPT-5 Mini: SI 50.6, released Aug 2025GPT-5 Nano: SI 39.1, released Aug 2025Qwen3 Max: SI 50.2, released Sep 2025Claude Sonnet 4.5: SI 53.9, released Sep 2025GLM-4.6: SI 47.8, released Sep 2025GPT-5 Pro: SI 45.8, released Oct 2025Claude Haiku 4.5: SI 60.5, released Oct 2025MiniMax-M2: SI 55.7, released Oct 2025Claude Opus 4.5: SI 62.6, released Nov 2025Kimi K2 Thinking Turbo: SI 66.1, released Nov 2025GPT-5.1: SI 55.7, released Nov 2025Gemini 3 Pro Preview: SI 58.6, released Nov 2025DeepSeek V3.2: SI 53.3, released Dec 2025GPT-5.2: SI 65.9, released Dec 2025Gemini 3 Flash Preview: SI 54.8, released Dec 2025GLM-4.7: SI 61.3, released Dec 2025Kimi K2.5: SI 59.0, released Jan 2026GLM-4.7-Flash: SI 53.3, released Jan 2026Step 3.5 Flash: SI 53.0, released Jan 2026Claude Opus 4.6: SI 70.8, released Feb 2026GLM-5: SI 59.4, released Feb 2026MiniMax-M2.5: SI 57.8, released Feb 2026Qwen3.5 397B-A17B: SI 64.0, released Feb 2026Claude Sonnet 4.6: SI 68.3, released Feb 2026Gemini 3.1 Pro Preview: SI 68.1, released Feb 2026Qwen3.5 35B-A3B: SI 62.7, released Feb 2026Qwen3.5 Flash: SI 59.5, released Feb 2026GPT-5.4: SI 69.7, released Mar 2026GPT-5.4 Pro: SI 61.3, released Mar 2026Grok 4.20 (Reasoning): SI 57.7, released Mar 2026GPT-5.4 nano: SI 57.4, released Mar 2026GPT-5.4 mini: SI 56.9, released Mar 2026MiniMax-M2.7: SI 57.9, released Mar 2026Gemma 4 26B A4B IT: SI 63.4, released Apr 2026Gemma 4 31B IT: SI 63.0, released Apr 2026Qwen3.6 Plus: SI 61.6, released Apr 2026GLM-5.1: SI 63.3, released Apr 2026Muse Spark 1.1: SI 62.9, released Apr 2026Claude Opus 4.7: SI 71.2, released Apr 2026Grok 4.3: SI 54.6, released Apr 2026Qwen3.6 Max Preview: SI 66.4, released Apr 2026Kimi K2.6: SI 64.0, released Apr 2026MiMo-V2.5-Pro: SI 62.6, released Apr 2026MiMo-V2.5: SI 56.4, released Apr 2026Qwen3.6 27B: SI 55.6, released Apr 2026GPT-5.5: SI 69.5, released Apr 2026GPT-5.5 Pro: SI 63.5, released Apr 2026DeepSeek V4 Pro: SI 63.1, released Apr 2026DeepSeek V4 Flash: SI 56.2, released Apr 2026GPT-5.5 Instant: SI 58.3, released May 2026Gemini 3.5 Flash: SI 68.3, released May 2026Qwen3.7 Max: SI 58.6, released May 2026Claude Opus 4.8: SI 67.5, released May 2026MiniMax-M3: SI 68.0, released Jun 2026Qwen3.7 Plus: SI 62.7, released Jun 2026Nemotron 3 Ultra 550B A55B: SI 61.2, released Jun 2026Claude Fable 5: SI 76.8, released Jun 2026GLM-5.2: SI 65.7, released Jun 2026Fugu Ultra: SI 57.9, released Jun 2026Fugu: SI 55.9, released Jun 2026Claude Sonnet 5: SI 67.7, released Jun 2026LongCat-2.0: SI 56.6, released Jun 2026Hy3: SI 60.7, released Jul 2026Grok 4.5: SI 64.8, released Jul 2026GPT-5.6 Sol: SI 72.1, released Jul 2026GPT-5.6 Terra: SI 68.2, released Jul 2026GPT-5.6 Luna: SI 63.4, released Jul 2026Inkling: SI 66.5, released Jul 2026Kimi K3: SI 71.0, released Jul 2026Gemini 3.6 Flash: SI 65.5, released Jul 2026Gemini 3.5 Flash Lite: SI 55.2, released Jul 2026Claude Opus 5: SI 74.9, released Jul 2026Inkling Small: SI 53.6, released Jul 2026DeepSeek V4 Flash 0731: SI 61.0, released Jul 2026Qwen3.8 Max: SI 68.7, released Aug 2026Muse Spark 1.2: SI 62.3, released Aug 2026Nemotron 3.5 Lightning 30B A3B: SI 46.1, released Aug 2026Grok 4.6: SI 66.6, released Aug 2026DeepSeek V4 Pro 0813: SI 63.0, released Aug 2026Gemini 3.7 Flash: SI 70.0, released Aug 2026GLM-5.3: SI 67.7, released Aug 2026Qwen3.8 27B: SI 61.9, released Aug 2026GLM-5.3-Flash: SI 66.2, released Aug 2026Claude Fable 5.1: SI 80.2, released Sep 2026Muse Spark 1.3: SI 70.1, released Sep 2026Gemini 3.8 Flash: SI 68.8, released Sep 2026GPT-6 Astra: SI 77.5, released Sep 2026DeepSeek V4.1 Flash: SI 66.2, released Sep 2026Grok 4.7: SI 65.5, released Sep 2026Claude Opus 5.5: SI 78.4, released Sep 2026GPT-6 Sol: SI 68.9, released Sep 2026GPT-6 Luna: SI 61.6, released Sep 2026MiMo-V2.6-Pro: SI 57.4, released Sep 2026MiMo-V2.6-Flash: SI 55.8, released Sep 2026Claude Sonnet 5.5: SI 69.4, released Sep 2026GPT-6.1 Sol: SI 73.0, released Sep 2026Mistral Large 4: SI 62.2, released Oct 2026Claude Haiku 5.5: SI 59.4, released Oct 2026 Claude Opus 4.6DeepSeek V3 0324GPT-4o (2024-05-13) Claude Fable 5.1 SI 80.2, Sep 2026 20 40 60 80 20252026 Claude Fable 5.1 80.2

What changed

Hear when the frontier moves

Get a notification on this device when a new model takes #1 or enters the top 10. No account, no email; turn it off any time.

What to be alerted about
RSS feed

Score against price

Best value models

Blended API price per million tokens, log scale. Up and to the left means more score per dollar.

20 40 60 80 100 $0.1$0.5$2.5$10$50 USD / 1M tokens (log scale; 3:1 in/out) Claude Fable 5.1 — SI 80.2 · $20.00/1M blended · confidence 100% Claude Opus 5.5 — SI 78.4 · $8.00/1M blended · confidence 100% GPT-6 Astra — SI 77.5 · $20.00/1M blended · confidence 100% Claude Fable 5 — SI 76.8 · $20.00/1M blended · confidence 100% Claude Opus 5 — SI 74.9 · $10.00/1M blended · confidence 100% GPT-6.1 Sol — SI 73.0 · $4.00/1M blended · confidence 100% GPT-5.6 Sol — SI 72.1 · $8.00/1M blended · confidence 100% Claude Opus 4.7 — SI 71.2 · $10.00/1M blended · confidence 100% Kimi K3 — SI 71.0 · $6.00/1M blended · confidence 100% Claude Opus 4.6 — SI 70.8 · $10.00/1M blended · confidence 100% Muse Spark 1.3 — SI 70.1 · $2.00/1M blended · confidence 93% Gemini 3.7 Flash — SI 70.0 · $1.50/1M blended · confidence 100% GPT-5.4 — SI 69.7 · $5.63/1M blended · confidence 100% GPT-5.5 — SI 69.5 · $11.25/1M blended · confidence 100% Claude Sonnet 5.5 — SI 69.4 · $4.00/1M blended · confidence 93% GPT-6 Sol — SI 68.9 · $4.00/1M blended · confidence 100% Gemini 3.8 Flash — SI 68.8 · $1.50/1M blended · confidence 100% Qwen3.8 Max — SI 68.7 · $2.48/1M blended · confidence 85% Gemini 3.5 Flash — SI 68.3 · $3.38/1M blended · confidence 100% Claude Sonnet 4.6 — SI 68.3 · $6.00/1M blended · confidence 100% GPT-5.6 Terra — SI 68.2 · $4.50/1M blended · confidence 100% Gemini 3.1 Pro Preview — SI 68.1 · $4.50/1M blended · confidence 100% MiniMax-M3 — SI 68.0 · $0.52/1M blended · confidence 100% GLM-5.3 — SI 67.7 · $2.15/1M blended · confidence 93% Claude Sonnet 5 — SI 67.7 · $4.00/1M blended · confidence 93% Claude Opus 4.8 — SI 67.5 · $10.00/1M blended · confidence 100% Grok 4.6 — SI 66.6 · $3.00/1M blended · confidence 100% Qwen3.6 Max Preview — SI 66.4 · $2.92/1M blended · confidence 64% GLM-5.3-Flash — SI 66.2 · $0.24/1M blended · confidence 100% DeepSeek V4.1 Flash — SI 66.2 · $0.26/1M blended · confidence 69% GPT-5.2 — SI 65.9 · $4.81/1M blended · confidence 100% GLM-5.2 — SI 65.7 · $2.15/1M blended · confidence 100% Gemini 3.6 Flash — SI 65.5 · $1.50/1M blended · confidence 100% Grok 4.7 — SI 65.5 · $3.00/1M blended · confidence 100% Grok 4.5 — SI 64.8 · $3.00/1M blended · confidence 100% Kimi K2.6 — SI 64.0 · $1.71/1M blended · confidence 85% Qwen3.5 397B-A17B — SI 64.0 · $0.39/1M blended · confidence 69% GPT-5.5 Pro — SI 63.5 · $67.50/1M blended · confidence 53% GPT-5.6 Luna — SI 63.4 · $0.45/1M blended · confidence 100% GLM-5.1 — SI 63.3 · $2.15/1M blended · confidence 64% DeepSeek V4 Pro — SI 63.1 · $1.98/1M blended · confidence 85% DeepSeek V4 Pro 0813 — SI 63.0 · $0.99/1M blended · confidence 69% Muse Spark 1.1 — SI 62.9 · $2.00/1M blended · confidence 61% Qwen3.7 Plus — SI 62.7 · $0.48/1M blended · confidence 64% Qwen3.5 35B-A3B — SI 62.7 · $0.16/1M blended · confidence 64% Claude Opus 4.5 — SI 62.6 · $10.00/1M blended · confidence 100% MiMo-V2.5-Pro — SI 62.6 · $0.54/1M blended · confidence 88% Muse Spark 1.2 — SI 62.3 · $2.00/1M blended · confidence 53% Mistral Large 4 — SI 62.2 · $1.03/1M blended · confidence 64% Qwen3.8 27B — SI 61.9 · $1.13/1M blended · confidence 69% GPT-6 Luna — SI 61.6 · $0.20/1M blended · confidence 100% Qwen3.6 Plus — SI 61.6 · $0.62/1M blended · confidence 80% GLM-4.7 — SI 61.3 · $1.00/1M blended · confidence 69% GPT-5.4 Pro — SI 61.3 · $67.50/1M blended · confidence 69% Claude Haiku 4.5 — SI 60.5 · $2.00/1M blended · confidence 64% Qwen3.5 Flash — SI 59.5 · $0.09/1M blended · confidence 64% Claude Haiku 5.5 — SI 59.4 · $0.20/1M blended · confidence 68% GLM-5 — SI 59.4 · $1.55/1M blended · confidence 90% Kimi K2.5 — SI 59.0 · $1.20/1M blended · confidence 100% Qwen3.7 Max — SI 58.6 · $2.48/1M blended · confidence 53% MiniMax-M2.7 — SI 57.9 · $0.52/1M blended · confidence 88% MiniMax-M2.5 — SI 57.8 · $0.52/1M blended · confidence 100% Grok 4.20 (Reasoning) — SI 57.7 · $1.56/1M blended · confidence 80% MiMo-V2.6-Pro — SI 57.4 · $0.54/1M blended · confidence 88% GPT-5.4 nano — SI 57.4 · $0.46/1M blended · confidence 100% GPT-5.4 mini — SI 56.9 · $1.69/1M blended · confidence 100% MiMo-V2.5 — SI 56.4 · $0.18/1M blended · confidence 88% DeepSeek V4 Flash — SI 56.2 · $0.52/1M blended · confidence 53% MiMo-V2.6-Flash — SI 55.8 · $0.18/1M blended · confidence 88% GPT-5 — SI 55.7 · $3.44/1M blended · confidence 100% MiniMax-M2 — SI 55.7 · $0.52/1M blended · confidence 96% GPT-5.1 — SI 55.7 · $3.44/1M blended · confidence 85% Qwen3.6 27B — SI 55.6 · $1.35/1M blended · confidence 53% Claude Opus 4 — SI 55.5 · $30.00/1M blended · confidence 100% o3 — SI 55.4 · $3.50/1M blended · confidence 87% Gemini 3.5 Flash Lite — SI 55.2 · $0.85/1M blended · confidence 100% Gemini 3 Flash Preview — SI 54.8 · $1.13/1M blended · confidence 93% Grok 4.3 — SI 54.6 · $1.56/1M blended · confidence 80% DeepSeek-V3 — SI 54.1 · $0.48/1M blended · confidence 93% Claude Sonnet 4.5 — SI 53.9 · $6.00/1M blended · confidence 80% DeepSeek V3.2 — SI 53.3 · $0.31/1M blended · confidence 56% Step 3.5 Flash — SI 53.0 · $0.15/1M blended · confidence 88% Qwen3 30B A3B — SI 52.6 · $0.19/1M blended · confidence 64% o4-mini — SI 51.2 · $1.93/1M blended · confidence 87% GPT-5 Mini — SI 50.6 · $0.69/1M blended · confidence 99% Qwen3 32B — SI 50.4 · $0.28/1M blended · confidence 67% Qwen3 235B-A22B — SI 50.3 · $0.50/1M blended · confidence 76% Qwen3 Max — SI 50.2 · $0.63/1M blended · confidence 69% GPT-4o (2024-05-13) — SI 49.4 · $7.50/1M blended · confidence 93% DeepSeek-R1 — SI 47.9 · $0.96/1M blended · confidence 88% GLM-4.6 — SI 47.8 · $1.00/1M blended · confidence 55% Claude Sonnet 4 — SI 46.6 · $6.00/1M blended · confidence 100% GPT-5 Pro — SI 45.8 · $41.25/1M blended · confidence 64% Gemini 2.5 Pro — SI 45.2 · $3.44/1M blended · confidence 90% GPT-4o (2024-11-20) — SI 42.2 · $4.38/1M blended · confidence 61% o3-mini — SI 41.6 · $1.93/1M blended · confidence 96% GPT-4.1 — SI 41.5 · $3.50/1M blended · confidence 100% Claude Haiku 3.5 — SI 40.8 · $1.60/1M blended · confidence 100% GPT-5 Nano — SI 39.1 · $0.14/1M blended · confidence 85% GPT-4.1 mini — SI 36.9 · $0.70/1M blended · confidence 87% GPT-4o (2024-08-06) — SI 36.7 · $4.38/1M blended · confidence 100% GPT-4.1 nano — SI 32.1 · $0.17/1M blended · confidence 82% GPT-4o — SI 31.2 · $4.38/1M blended · confidence 59% GPT-4o mini — SI 22.9 · $0.26/1M blended · confidence 100% Claude Fable 5.1GPT-6.1 SolGemini 3.7 FlashMiniMax-M3Qwen3.5 35B-A3B 20 40 60 80 100 $0.1$0.5$2.5$10$50 USD / 1M tokens (log scale; 3:1 in/out) Claude Fable 5.1 — SI 80.2 · $20.00/1M blended · confidence 100% Claude Opus 5.5 — SI 78.4 · $8.00/1M blended · confidence 100% GPT-6 Astra — SI 77.5 · $20.00/1M blended · confidence 100% Claude Fable 5 — SI 76.8 · $20.00/1M blended · confidence 100% Claude Opus 5 — SI 74.9 · $10.00/1M blended · confidence 100% GPT-6.1 Sol — SI 73.0 · $4.00/1M blended · confidence 100% GPT-5.6 Sol — SI 72.1 · $8.00/1M blended · confidence 100% Claude Opus 4.7 — SI 71.2 · $10.00/1M blended · confidence 100% Kimi K3 — SI 71.0 · $6.00/1M blended · confidence 100% Claude Opus 4.6 — SI 70.8 · $10.00/1M blended · confidence 100% Muse Spark 1.3 — SI 70.1 · $2.00/1M blended · confidence 93% Gemini 3.7 Flash — SI 70.0 · $1.50/1M blended · confidence 100% GPT-5.4 — SI 69.7 · $5.63/1M blended · confidence 100% GPT-5.5 — SI 69.5 · $11.25/1M blended · confidence 100% Claude Sonnet 5.5 — SI 69.4 · $4.00/1M blended · confidence 93% GPT-6 Sol — SI 68.9 · $4.00/1M blended · confidence 100% Gemini 3.8 Flash — SI 68.8 · $1.50/1M blended · confidence 100% Qwen3.8 Max — SI 68.7 · $2.48/1M blended · confidence 85% Gemini 3.5 Flash — SI 68.3 · $3.38/1M blended · confidence 100% Claude Sonnet 4.6 — SI 68.3 · $6.00/1M blended · confidence 100% GPT-5.6 Terra — SI 68.2 · $4.50/1M blended · confidence 100% Gemini 3.1 Pro Preview — SI 68.1 · $4.50/1M blended · confidence 100% MiniMax-M3 — SI 68.0 · $0.52/1M blended · confidence 100% GLM-5.3 — SI 67.7 · $2.15/1M blended · confidence 93% Claude Sonnet 5 — SI 67.7 · $4.00/1M blended · confidence 93% Claude Opus 4.8 — SI 67.5 · $10.00/1M blended · confidence 100% Grok 4.6 — SI 66.6 · $3.00/1M blended · confidence 100% Qwen3.6 Max Preview — SI 66.4 · $2.92/1M blended · confidence 64% GLM-5.3-Flash — SI 66.2 · $0.24/1M blended · confidence 100% DeepSeek V4.1 Flash — SI 66.2 · $0.26/1M blended · confidence 69% GPT-5.2 — SI 65.9 · $4.81/1M blended · confidence 100% GLM-5.2 — SI 65.7 · $2.15/1M blended · confidence 100% Gemini 3.6 Flash — SI 65.5 · $1.50/1M blended · confidence 100% Grok 4.7 — SI 65.5 · $3.00/1M blended · confidence 100% Grok 4.5 — SI 64.8 · $3.00/1M blended · confidence 100% Kimi K2.6 — SI 64.0 · $1.71/1M blended · confidence 85% Qwen3.5 397B-A17B — SI 64.0 · $0.39/1M blended · confidence 69% GPT-5.5 Pro — SI 63.5 · $67.50/1M blended · confidence 53% GPT-5.6 Luna — SI 63.4 · $0.45/1M blended · confidence 100% GLM-5.1 — SI 63.3 · $2.15/1M blended · confidence 64% DeepSeek V4 Pro — SI 63.1 · $1.98/1M blended · confidence 85% DeepSeek V4 Pro 0813 — SI 63.0 · $0.99/1M blended · confidence 69% Muse Spark 1.1 — SI 62.9 · $2.00/1M blended · confidence 61% Qwen3.7 Plus — SI 62.7 · $0.48/1M blended · confidence 64% Qwen3.5 35B-A3B — SI 62.7 · $0.16/1M blended · confidence 64% Claude Opus 4.5 — SI 62.6 · $10.00/1M blended · confidence 100% MiMo-V2.5-Pro — SI 62.6 · $0.54/1M blended · confidence 88% Muse Spark 1.2 — SI 62.3 · $2.00/1M blended · confidence 53% Mistral Large 4 — SI 62.2 · $1.03/1M blended · confidence 64% Qwen3.8 27B — SI 61.9 · $1.13/1M blended · confidence 69% GPT-6 Luna — SI 61.6 · $0.20/1M blended · confidence 100% Qwen3.6 Plus — SI 61.6 · $0.62/1M blended · confidence 80% GLM-4.7 — SI 61.3 · $1.00/1M blended · confidence 69% GPT-5.4 Pro — SI 61.3 · $67.50/1M blended · confidence 69% Claude Haiku 4.5 — SI 60.5 · $2.00/1M blended · confidence 64% Qwen3.5 Flash — SI 59.5 · $0.09/1M blended · confidence 64% Claude Haiku 5.5 — SI 59.4 · $0.20/1M blended · confidence 68% GLM-5 — SI 59.4 · $1.55/1M blended · confidence 90% Kimi K2.5 — SI 59.0 · $1.20/1M blended · confidence 100% Qwen3.7 Max — SI 58.6 · $2.48/1M blended · confidence 53% MiniMax-M2.7 — SI 57.9 · $0.52/1M blended · confidence 88% MiniMax-M2.5 — SI 57.8 · $0.52/1M blended · confidence 100% Grok 4.20 (Reasoning) — SI 57.7 · $1.56/1M blended · confidence 80% MiMo-V2.6-Pro — SI 57.4 · $0.54/1M blended · confidence 88% GPT-5.4 nano — SI 57.4 · $0.46/1M blended · confidence 100% GPT-5.4 mini — SI 56.9 · $1.69/1M blended · confidence 100% MiMo-V2.5 — SI 56.4 · $0.18/1M blended · confidence 88% DeepSeek V4 Flash — SI 56.2 · $0.52/1M blended · confidence 53% MiMo-V2.6-Flash — SI 55.8 · $0.18/1M blended · confidence 88% GPT-5 — SI 55.7 · $3.44/1M blended · confidence 100% MiniMax-M2 — SI 55.7 · $0.52/1M blended · confidence 96% GPT-5.1 — SI 55.7 · $3.44/1M blended · confidence 85% Qwen3.6 27B — SI 55.6 · $1.35/1M blended · confidence 53% Claude Opus 4 — SI 55.5 · $30.00/1M blended · confidence 100% o3 — SI 55.4 · $3.50/1M blended · confidence 87% Gemini 3.5 Flash Lite — SI 55.2 · $0.85/1M blended · confidence 100% Gemini 3 Flash Preview — SI 54.8 · $1.13/1M blended · confidence 93% Grok 4.3 — SI 54.6 · $1.56/1M blended · confidence 80% DeepSeek-V3 — SI 54.1 · $0.48/1M blended · confidence 93% Claude Sonnet 4.5 — SI 53.9 · $6.00/1M blended · confidence 80% DeepSeek V3.2 — SI 53.3 · $0.31/1M blended · confidence 56% Step 3.5 Flash — SI 53.0 · $0.15/1M blended · confidence 88% Qwen3 30B A3B — SI 52.6 · $0.19/1M blended · confidence 64% o4-mini — SI 51.2 · $1.93/1M blended · confidence 87% GPT-5 Mini — SI 50.6 · $0.69/1M blended · confidence 99% Qwen3 32B — SI 50.4 · $0.28/1M blended · confidence 67% Qwen3 235B-A22B — SI 50.3 · $0.50/1M blended · confidence 76% Qwen3 Max — SI 50.2 · $0.63/1M blended · confidence 69% GPT-4o (2024-05-13) — SI 49.4 · $7.50/1M blended · confidence 93% DeepSeek-R1 — SI 47.9 · $0.96/1M blended · confidence 88% GLM-4.6 — SI 47.8 · $1.00/1M blended · confidence 55% Claude Sonnet 4 — SI 46.6 · $6.00/1M blended · confidence 100% GPT-5 Pro — SI 45.8 · $41.25/1M blended · confidence 64% Gemini 2.5 Pro — SI 45.2 · $3.44/1M blended · confidence 90% GPT-4o (2024-11-20) — SI 42.2 · $4.38/1M blended · confidence 61% o3-mini — SI 41.6 · $1.93/1M blended · confidence 96% GPT-4.1 — SI 41.5 · $3.50/1M blended · confidence 100% Claude Haiku 3.5 — SI 40.8 · $1.60/1M blended · confidence 100% GPT-5 Nano — SI 39.1 · $0.14/1M blended · confidence 85% GPT-4.1 mini — SI 36.9 · $0.70/1M blended · confidence 87% GPT-4o (2024-08-06) — SI 36.7 · $4.38/1M blended · confidence 100% GPT-4.1 nano — SI 32.1 · $0.17/1M blended · confidence 82% GPT-4o — SI 31.2 · $4.38/1M blended · confidence 59% GPT-4o mini — SI 22.9 · $0.26/1M blended · confidence 100%
Best score at its price or lower Other ranked models (fainter means lower confidence) Open weights

Latest releases

All releases

Understand the numbers

What changed

All releases · RSS feed

Alerts on this device

What to be alerted about
RSS feed