SI Score · updated Oct 9, 2026

Superintelligence leaderboard.
Open evidence. Clear rankings.

The latest frontier models, ranked by the SI Score — a composite of open benchmarks, scaled per benchmark and weighted across reasoning, coding, math and human preference. Every number links to its source and date, and each score carries a confidence % that rises as sources report.

449
models tracked
58
benchmarks in the composite
29
sources tracked
80.2
top SI Score — Claude Fable 5.1

The leaderboard

Full catalog →
1 Claude Fable 5.1 Anthropic 80.2 100% confidence 100 percent, Full $10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$50.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 1, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
2 Claude Opus 5.5 Anthropic 78.4 100% confidence 100 percent, Full $4.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$20.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 22, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
3 GPT-6 Astra OpenAI 77.5 100% confidence 100 percent, Full $10.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
$50.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 3, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
4 Claude Fable 5 Anthropic 76.8 100% confidence 100 percent, Full $10.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$50.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
5 Claude Opus 5 Anthropic 74.9 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 24, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
6 GPT-6.1 Sol OpenAI 73.0 100% confidence 100 percent, Full $2.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
$10.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 29, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
7 GPT-5.6 Sol OpenAI 72.1 100% confidence 100 percent, Full $4.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
$20.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
8 Claude Opus 4.7 Anthropic 71.2 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 16, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
9 Kimi K3 Moonshot AI 71.0 100% confidence 100 percent, Full $3.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k3 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 16, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
10 Claude Opus 4.6 Anthropic 70.8 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 5, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
11 Muse Spark 1.3 Meta 70.1 93% confidence 93 percent, High $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
12 Gemini 3.7 Flash Google 70.0 100% confidence 100 percent, Full $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 13, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
13 GPT-5.4 OpenAI 69.7 100% confidence 100 percent, Full $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4 Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4 Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 5, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
14 GPT-5.5 OpenAI 69.5 100% confidence 100 percent, Full $5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
15 Claude Sonnet 5.5 Anthropic 69.4 93% confidence 93 percent, High $2.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 28, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
16 GPT-6 Sol OpenAI 68.9 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-6-sol Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
17 Gemini 3.8 Flash Google 68.8 100% confidence 100 percent, Full $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 2, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
18 Qwen3.8 Max Alibaba / Qwen 68.7 85% confidence 85 percent, High $1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.95Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 3, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
19 Gemini 3.5 Flash Google 68.3 100% confidence 100 percent, Full $1.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$9.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 19, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
20 Claude Sonnet 4.6 Anthropic 68.3 100% confidence 100 percent, Full $3.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 17, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
21 GPT-5.6 Terra OpenAI 68.2 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-terra Retrieved Oct 9, 2026 · MIT
Open source ↗
$12.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-terra Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
22 Gemini 3.1 Pro Preview Google 68.1 100% confidence 100 percent, Full $2.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$12.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 19, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
23 MiniMax-M3 minimax 68.0 100% confidence 100 percent, Full $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 1, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
24 GLM-5.3 Z.ai 67.7 93% confidence 93 percent, High $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 14, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
25 Claude Sonnet 5 Anthropic 67.7 93% confidence 93 percent, High $2.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$10.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 30, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
26 Claude Opus 4.8 Anthropic 67.5 100% confidence 100 percent, Full $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$25.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 28, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
27 Grok 4.6 xAI 66.6 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
$6.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
500Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
28 Inkling thinkingmachines 66.5 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
29 Qwen3.6 Max Preview Alibaba / Qwen 66.4 64% confidence 64 percent, Medium $1.30Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤128K; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$7.80Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤128K; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 20, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
30 GLM-5.3-Flash Z.ai 66.2 100% confidence 100 percent, Full $0.15Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.50Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 26, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
31 DeepSeek V4.1 Flash DeepSeek 66.2 69% confidence 69 percent, Medium $0.15DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.60DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 10, 2026DeepSeek V4.1 Flash announcementOfficial dated introduction and availability announcement Retrieved Oct 9, 2026 · factual citation
Open source ↗
Open
32 Kimi K2 Thinking Turbo Moonshot AI 66.1 64% confidence 64 percent, Medium not yet reported not yet reported 262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 6, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
33 GPT-5.2 OpenAI 65.9 100% confidence 100 percent, Full $1.75models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.2 Retrieved Oct 9, 2026 · MIT
Open source ↗
$14.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.2 Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 11, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
34 GLM-5.2 Z.ai 65.7 100% confidence 100 percent, Full $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 13, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
35 Gemini 3.6 Flash Google 65.5 100% confidence 100 percent, Full $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 21, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
36 Grok 4.7 xAI 65.5 100% confidence 100 percent, Full $2.00xAI models & pricingPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
$6.00xAI models & pricingPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
500KxAI models & pricingPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
Sep 21, 2026xAI models & pricingOfficial featured model page datePublished; model introduction date Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
37 Grok 4.5 xAI 64.8 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$6.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
500Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 8, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
38 Kimi K2.6 Moonshot AI 64.0 85% confidence 85 percent, High $0.95models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k2.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k2.6 Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 21, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
39 Qwen3.5 397B-A17B Alibaba / Qwen 64.0 69% confidence 69 percent, Medium $0.17Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.03Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
40 GPT-5.5 Pro OpenAI 63.5 53% confidence 53 percent, Medium $30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$180.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
41 GPT-5.6 Luna OpenAI 63.4 100% confidence 100 percent, Full $0.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-luna Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-luna Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
42 Gemma 4 26B A4B IT Google 63.4 64% confidence 64 percent, Medium $0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
43 GLM-5.1 Z.ai 63.3 64% confidence 64 percent, Medium $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 7, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
44 DeepSeek V4 Pro DeepSeek 63.1 85% confidence 85 percent, High $1.32LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$3.96LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
45 Gemma 4 31B IT Google 63.0 64% confidence 64 percent, Medium $0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
46 DeepSeek V4 Pro 0813 DeepSeek 63.0 69% confidence 69 percent, Medium $0.66DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.98DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
47 Muse Spark 1.1 Meta 62.9 61% confidence 61 percent, Medium $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 8, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
48 Qwen3.7 Plus Alibaba / Qwen 62.7 64% confidence 64 percent, Medium $0.28Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.10Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
49 Qwen3.5 35B-A3B Alibaba / Qwen 62.7 64% confidence 64 percent, Medium $0.06Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.46Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 23, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
50 Claude Opus 4.5 Anthropic 62.6 100% confidence 100 percent, Full $5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-opus-4-5-20251101 Retrieved Oct 9, 2026 · MIT
Open source ↗
$25.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-opus-4-5-20251101 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
51 MiMo-V2.5-Pro xiaomi 62.6 88% confidence 88 percent, High $0.43models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.87models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
52 Muse Spark 1.2 Meta 62.3 53% confidence 53 percent, Medium $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
53 Mistral Large 4 mistral 62.2 64% confidence 64 percent, Medium $0.68Mistral API pricingOfficial Mistral Serverless API Standard rate, displayed sale price where applicable; cache, Batch, specialist units and hosted third-party models excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$2.09Mistral API pricingOfficial Mistral Serverless API Standard rate, displayed sale price where applicable; cache, Batch, specialist units and hosted third-party models excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 6, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
54 Qwen3.8 27B Alibaba / Qwen 61.9 69% confidence 69 percent, Medium $0.50Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$3.00Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 14, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
55 GPT-6 Luna OpenAI 61.6 100% confidence 100 percent, Full $0.10OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.50OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts Retrieved Oct 9, 2026 · factual citation
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
56 Qwen3.6 Plus Alibaba / Qwen 61.6 80% confidence 80 percent, High $0.28Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 2, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
57 GLM-4.7 Z.ai 61.3 69% confidence 69 percent, Medium $0.60Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$2.20Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 22, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
58 GPT-5.4 Pro OpenAI 61.3 69% confidence 69 percent, Medium $30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$180.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1.1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 5, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
59 Nemotron 3 Ultra 550B A55B nvidia 61.2 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 4, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
60 DeepSeek V4 Flash 0731 DeepSeek 61.0 69% confidence 69 percent, Medium not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 31, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
61 Hy3 tencent 60.7 88% confidence 88 percent, High $0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://cloud.tencent.com/document/product/1823/130050. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; tencent-tokenhub/hy3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://cloud.tencent.com/document/product/1823/130050. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; tencent-tokenhub/hy3 Retrieved Oct 9, 2026 · MIT
Open source ↗
256Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 6, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
62 Claude Haiku 4.5 Anthropic 60.5 64% confidence 64 percent, Medium $1.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-haiku-4-5-20251001 Retrieved Oct 9, 2026 · MIT
Open source ↗
$5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-haiku-4-5-20251001 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 15, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
63 Qwen3.5 Flash Alibaba / Qwen 59.5 64% confidence 64 percent, Medium $0.03Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.29Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 23, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
64 Claude Haiku 5.5 Anthropic 59.4 68% confidence 68 percent, Medium $0.10Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.50Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
Oct 7, 2026Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
65 GLM-5 Z.ai 59.4 90% confidence 90 percent, High $1.00Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$3.20Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
66 Kimi K2.5 Moonshot AI 59.0 100% confidence 100 percent, Full $0.60LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://platform.moonshot.ai/docs/guide/kimi-k2-5-quickstart. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$3.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://platform.moonshot.ai/docs/guide/kimi-k2-5-quickstart. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 1, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
67 Qwen3.7 Max Alibaba / Qwen 58.6 53% confidence 53 percent, Medium $1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.95Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 21, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
68 Gemini 3 Pro Preview Google 58.6 68% confidence 68 percent, Medium not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 18, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
69 GPT-5.5 Instant OpenAI 58.3 64% confidence 64 percent, Medium not yet reported not yet reported 400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 5, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
70 MiniMax-M2.7 minimax 57.9 88% confidence 88 percent, High $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 18, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
71 Fugu Ultra sakana 57.9 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
72 MiniMax-M2.5 minimax 57.8 100% confidence 100 percent, Full $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 12, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
73 Grok 4.20 (Reasoning) xAI 57.7 80% confidence 80 percent, High $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.20-0309-reasoning Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.20-0309-reasoning Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 9, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
74 MiMo-V2.6-Pro xiaomi 57.4 88% confidence 88 percent, High $0.43models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.87models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
75 GPT-5.4 nano OpenAI 57.4 100% confidence 100 percent, Full $0.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 17, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
76 GPT-5.4 mini OpenAI 56.9 100% confidence 100 percent, Full $0.75models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 17, 2026OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
77 LongCat-2.0 meituan 56.6 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 30, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
78 MiMo-V2.5 xiaomi 56.4 88% confidence 88 percent, High $0.14models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.28models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
79 DeepSeek V4 Flash DeepSeek 56.2 53% confidence 53 percent, Medium $0.30LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.20LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 24, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
80 Fugu sakana 55.9 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 15, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
81 MiMo-V2.6-Flash xiaomi 55.8 88% confidence 88 percent, High $0.14models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.28models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
82 DeepSeek V3 0324 DeepSeek 55.7 74% confidence 74 percent, Medium not yet reported not yet reported 164Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 24, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
83 GPT-5 OpenAI 55.7 100% confidence 100 percent, Full $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5 Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 7, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
84 MiniMax-M2 minimax 55.7 96% confidence 96 percent, High $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.20MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 27, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
85 GPT-5.1 OpenAI 55.7 85% confidence 85 percent, High $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 13, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
86 Qwen3.6 27B Alibaba / Qwen 55.6 53% confidence 53 percent, Medium $0.60Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$3.60Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 22, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
87 Claude Opus 4 Anthropic 55.5 100% confidence 100 percent, Full $15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$75.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 22, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
88 o3 OpenAI 55.4 87% confidence 87 percent, High $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/o3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$8.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/o3 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 16, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
89 Gemini 3.5 Flash Lite Google 55.2 100% confidence 100 percent, Full $0.30Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$2.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 21, 2026Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
90 GPT OSS 120B OpenAI 54.8 79% confidence 79 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
91 Gemini 3 Flash Preview Google 54.8 93% confidence 93 percent, High $0.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$3.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 17, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
92 Grok 4.3 xAI 54.6 80% confidence 80 percent, High $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.3 Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.3 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 17, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
93 DeepSeek-V3 DeepSeek 54.1 93% confidence 93 percent, High $0.27LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 26, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
94 Claude Sonnet 4.5 Anthropic 53.9 80% confidence 80 percent, High $3.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-sonnet-4-5-20250929 Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-sonnet-4-5-20250929 Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 29, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
95 Inkling Small thinkingmachines 53.6 100% confidence 100 percent, Full not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 30, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
96 DeepSeek V3.2 DeepSeek 53.3 56% confidence 56 percent, Medium $0.28LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
97 GLM-4.7-Flash Z.ai 53.3 69% confidence 69 percent, Medium $0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.z.ai/guides/overview/pricing. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; zai/glm-4.7-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.z.ai/guides/overview/pricing. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; zai/glm-4.7-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 19, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
98 Step 3.5 Flash stepfun 53.0 88% confidence 88 percent, High $0.10models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.stepfun.com/docs/zh/overview/concept. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; stepfun/step-3.5-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.30models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.stepfun.com/docs/zh/overview/concept. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; stepfun/step-3.5-flash Retrieved Oct 9, 2026 · MIT
Open source ↗
256Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 29, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
99 Qwen3 30B A3B Alibaba / Qwen 52.6 64% confidence 64 percent, Medium $0.11Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.43Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 28, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
100 GPT OSS 20B OpenAI 52.5 64% confidence 64 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
101 Claude Sonnet 3.5 v2 Anthropic 51.9 100% confidence 100 percent, Full not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 22, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
102 Claude Opus 4.1 Anthropic 51.5 80% confidence 80 percent, High not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
103 o4-mini OpenAI 51.2 87% confidence 87 percent, High $1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 16, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
104 Llama 4 Scout 17B Instruct Meta 51.0 64% confidence 64 percent, Medium not yet reported not yet reported 10Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
105 GPT-5 Mini OpenAI 50.6 99% confidence 99 percent, High $0.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 7, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
106 Qwen3 32B Alibaba / Qwen 50.4 67% confidence 67 percent, Medium $0.16Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$0.64Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
107 Qwen3 235B-A22B Alibaba / Qwen 50.3 76% confidence 76 percent, Medium $0.29Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.15Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 1, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
108 Qwen3 Max Alibaba / Qwen 50.2 69% confidence 69 percent, Medium $0.36Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤32K; Non-Thinking mode only. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$1.43Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤32K; Non-Thinking mode only. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 23, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
109 GPT-4o (2024-05-13) OpenAI 49.4 93% confidence 93 percent, High $5.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$15.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 13, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
110 Claude Sonnet 3.7 Anthropic 48.9 100% confidence 100 percent, Full not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Feb 19, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
111 Mistral Small 3.1 24B mistral 48.4 64% confidence 64 percent, Medium not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 17, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
112 QwQ 32B Alibaba / Qwen 48.2 67% confidence 67 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
113 DeepSeek-R1 DeepSeek 47.9 88% confidence 88 percent, High $0.55LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$2.19LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jan 20, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
114 GLM-4.6 Z.ai 47.8 55% confidence 55 percent, Medium $0.60Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$2.20Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
205Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 30, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
115 Claude Sonnet 4 Anthropic 46.6 100% confidence 100 percent, Full $3.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 22, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
116 Nemotron 3.5 Lightning 30B A3B nvidia 46.1 100% confidence 100 percent, Full not yet reported not yet reported 262Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 11, 2026models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
117 GPT-5 Pro OpenAI 45.8 64% confidence 64 percent, Medium $15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
$120.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-pro Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 6, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
118 Llama-3.1-70B-Instruct Meta 45.7 93% confidence 93 percent, High not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 23, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
119 Gemini 2.5 Pro Google 45.2 90% confidence 90 percent, High $1.25Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
$10.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jun 5, 2025Google Gemini release notesPublished source fact Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation
Open source ↗
—
120 Gemma 3 12B IT Google 45.1 64% confidence 64 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 12, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
121 GPT-4o (2024-11-20) OpenAI 42.2 61% confidence 61 percent, Medium $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-11-20 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-11-20 Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 20, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
122 Gemma 3 27B IT Google 42.0 74% confidence 74 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 12, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
123 o3-mini OpenAI 41.6 96% confidence 96 percent, High $1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$4.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 20, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
124 GPT-4.1 OpenAI 41.5 100% confidence 100 percent, Full $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
$8.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1 Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 14, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
125 Claude Haiku 3.5 Anthropic 40.8 100% confidence 100 percent, Full $0.80Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
$4.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded Retrieved Oct 9, 2026 · factual citation
Open source ↗
200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Oct 22, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
126 Claude Haiku 3 Anthropic 39.9 93% confidence 93 percent, High not yet reported not yet reported 200Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 13, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
127 Qwen2.5-Coder-32B-Instruct Alibaba / Qwen 39.6 90% confidence 90 percent, High not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 12, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
128 Gemma 3 4B IT Google 39.2 64% confidence 64 percent, Medium not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Mar 12, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
129 GPT-5 Nano OpenAI 39.1 85% confidence 85 percent, High $0.05models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.40models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-nano Retrieved Oct 9, 2026 · MIT
Open source ↗
400Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 7, 2025OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
130 Llama-3.1-8B-Instruct Meta 37.7 93% confidence 93 percent, High not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 23, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
131 Nova Pro amazon 37.4 100% confidence 100 percent, Full not yet reported not yet reported 300Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 3, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
132 GPT-4.1 mini OpenAI 36.9 87% confidence 87 percent, High $0.40models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$1.60models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 14, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
133 GPT-4o (2024-08-06) OpenAI 36.7 100% confidence 100 percent, Full $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-08-06 Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-08-06 Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Aug 6, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
134 Nova Lite amazon 36.6 100% confidence 100 percent, Full not yet reported not yet reported 300Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 3, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
135 Mistral Large 2.1 mistral 35.3 100% confidence 100 percent, Full not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Nov 18, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
136 Mistral Medium 3 mistral 33.0 85% confidence 85 percent, High not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 7, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
137 Llama-3.3-70B-Instruct Meta 32.7 100% confidence 100 percent, Full not yet reported not yet reported 128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Dec 6, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
138 Llama-3.2-1B Meta 32.1 93% confidence 93 percent, High not yet reported not yet reported 131Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Sep 25, 2024models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
139 GPT-4.1 nano OpenAI 32.1 82% confidence 82 percent, High $0.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.40LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded Retrieved Oct 9, 2026 · MIT
Open source ↗
1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 14, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
—
140 GPT-4o OpenAI 31.2 59% confidence 59 percent, Medium $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o Retrieved Oct 9, 2026 · MIT
Open source ↗
$10.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
May 13, 2024OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—
141 Llama 4 Maverick 17B Instruct Meta 30.0 78% confidence 78 percent, Medium not yet reported not yet reported 1Mmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Apr 5, 2025models.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Open
142 GPT-4o mini OpenAI 22.9 100% confidence 100 percent, Full $0.15models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
$0.60models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-mini Retrieved Oct 9, 2026 · MIT
Open source ↗
128Kmodels.devPublished source fact Retrieved Oct 9, 2026 · MIT
Open source ↗
Jul 18, 2024OpenAI API changelogPublished source fact Retrieved Oct 9, 2026 · factual citation
Open source ↗
—

Ranked text models appear first. The provisional toggle includes only text models awaiting evidence. Browse every kind in the model catalog. Prices are USD per 1M tokens (official first-party API). Dotted values carry their source — hover or tap to see it. “not yet reported” means the source has not published that figure for this model.

Score against price, and over time

Best value models →
SI Score 20 40 60 80 100 $0.1 $0.5 $2.5 $10 $50 USD / 1M tokens (log scale; 3:1 in/out) Claude Fable 5.1 — SI 80.2 · $20.00/1M blended · confidence 100% Claude Opus 5.5 — SI 78.4 · $8.00/1M blended · confidence 100% GPT-6 Astra — SI 77.5 · $20.00/1M blended · confidence 100% Claude Fable 5 — SI 76.8 · $20.00/1M blended · confidence 100% Claude Opus 5 — SI 74.9 · $10.00/1M blended · confidence 100% GPT-6.1 Sol — SI 73.0 · $4.00/1M blended · confidence 100% GPT-5.6 Sol — SI 72.1 · $8.00/1M blended · confidence 100% Claude Opus 4.7 — SI 71.2 · $10.00/1M blended · confidence 100% Kimi K3 — SI 71.0 · $6.00/1M blended · confidence 100% Claude Opus 4.6 — SI 70.8 · $10.00/1M blended · confidence 100% Muse Spark 1.3 — SI 70.1 · $2.00/1M blended · confidence 93% Gemini 3.7 Flash — SI 70.0 · $1.50/1M blended · confidence 100% GPT-5.4 — SI 69.7 · $5.63/1M blended · confidence 100% GPT-5.5 — SI 69.5 · $11.25/1M blended · confidence 100% Claude Sonnet 5.5 — SI 69.4 · $4.00/1M blended · confidence 93% GPT-6 Sol — SI 68.9 · $4.00/1M blended · confidence 100% Gemini 3.8 Flash — SI 68.8 · $1.50/1M blended · confidence 100% Qwen3.8 Max — SI 68.7 · $2.48/1M blended · confidence 85% Gemini 3.5 Flash — SI 68.3 · $3.38/1M blended · confidence 100% Claude Sonnet 4.6 — SI 68.3 · $6.00/1M blended · confidence 100% GPT-5.6 Terra — SI 68.2 · $4.50/1M blended · confidence 100% Gemini 3.1 Pro Preview — SI 68.1 · $4.50/1M blended · confidence 100% MiniMax-M3 — SI 68.0 · $0.52/1M blended · confidence 100% GLM-5.3 — SI 67.7 · $2.15/1M blended · confidence 93% Claude Sonnet 5 — SI 67.7 · $4.00/1M blended · confidence 93% Claude Opus 4.8 — SI 67.5 · $10.00/1M blended · confidence 100% Grok 4.6 — SI 66.6 · $3.00/1M blended · confidence 100% Qwen3.6 Max Preview — SI 66.4 · $2.92/1M blended · confidence 64% GLM-5.3-Flash — SI 66.2 · $0.24/1M blended · confidence 100% DeepSeek V4.1 Flash — SI 66.2 · $0.26/1M blended · confidence 69% GPT-5.2 — SI 65.9 · $4.81/1M blended · confidence 100% GLM-5.2 — SI 65.7 · $2.15/1M blended · confidence 100% Gemini 3.6 Flash — SI 65.5 · $1.50/1M blended · confidence 100% Grok 4.7 — SI 65.5 · $3.00/1M blended · confidence 100% Grok 4.5 — SI 64.8 · $3.00/1M blended · confidence 100% Kimi K2.6 — SI 64.0 · $1.71/1M blended · confidence 85% Qwen3.5 397B-A17B — SI 64.0 · $0.39/1M blended · confidence 69% GPT-5.5 Pro — SI 63.5 · $67.50/1M blended · confidence 53% GPT-5.6 Luna — SI 63.4 · $0.45/1M blended · confidence 100% GLM-5.1 — SI 63.3 · $2.15/1M blended · confidence 64% DeepSeek V4 Pro — SI 63.1 · $1.98/1M blended · confidence 85% DeepSeek V4 Pro 0813 — SI 63.0 · $0.99/1M blended · confidence 69% Muse Spark 1.1 — SI 62.9 · $2.00/1M blended · confidence 61% Qwen3.7 Plus — SI 62.7 · $0.48/1M blended · confidence 64% Qwen3.5 35B-A3B — SI 62.7 · $0.16/1M blended · confidence 64% Claude Opus 4.5 — SI 62.6 · $10.00/1M blended · confidence 100% MiMo-V2.5-Pro — SI 62.6 · $0.54/1M blended · confidence 88% Muse Spark 1.2 — SI 62.3 · $2.00/1M blended · confidence 53% Mistral Large 4 — SI 62.2 · $1.03/1M blended · confidence 64% Qwen3.8 27B — SI 61.9 · $1.13/1M blended · confidence 69% GPT-6 Luna — SI 61.6 · $0.20/1M blended · confidence 100% Qwen3.6 Plus — SI 61.6 · $0.62/1M blended · confidence 80% GLM-4.7 — SI 61.3 · $1.00/1M blended · confidence 69% GPT-5.4 Pro — SI 61.3 · $67.50/1M blended · confidence 69% Claude Haiku 4.5 — SI 60.5 · $2.00/1M blended · confidence 64% Qwen3.5 Flash — SI 59.5 · $0.09/1M blended · confidence 64% Claude Haiku 5.5 — SI 59.4 · $0.20/1M blended · confidence 68% GLM-5 — SI 59.4 · $1.55/1M blended · confidence 90% Kimi K2.5 — SI 59.0 · $1.20/1M blended · confidence 100% Qwen3.7 Max — SI 58.6 · $2.48/1M blended · confidence 53% MiniMax-M2.7 — SI 57.9 · $0.52/1M blended · confidence 88% MiniMax-M2.5 — SI 57.8 · $0.52/1M blended · confidence 100% Grok 4.20 (Reasoning) — SI 57.7 · $1.56/1M blended · confidence 80% MiMo-V2.6-Pro — SI 57.4 · $0.54/1M blended · confidence 88% GPT-5.4 nano — SI 57.4 · $0.46/1M blended · confidence 100% GPT-5.4 mini — SI 56.9 · $1.69/1M blended · confidence 100% MiMo-V2.5 — SI 56.4 · $0.18/1M blended · confidence 88% DeepSeek V4 Flash — SI 56.2 · $0.52/1M blended · confidence 53% MiMo-V2.6-Flash — SI 55.8 · $0.18/1M blended · confidence 88% GPT-5 — SI 55.7 · $3.44/1M blended · confidence 100% MiniMax-M2 — SI 55.7 · $0.52/1M blended · confidence 96% GPT-5.1 — SI 55.7 · $3.44/1M blended · confidence 85% Qwen3.6 27B — SI 55.6 · $1.35/1M blended · confidence 53% Claude Opus 4 — SI 55.5 · $30.00/1M blended · confidence 100% o3 — SI 55.4 · $3.50/1M blended · confidence 87% Gemini 3.5 Flash Lite — SI 55.2 · $0.85/1M blended · confidence 100% Gemini 3 Flash Preview — SI 54.8 · $1.13/1M blended · confidence 93% Grok 4.3 — SI 54.6 · $1.56/1M blended · confidence 80% DeepSeek-V3 — SI 54.1 · $0.48/1M blended · confidence 93% Claude Sonnet 4.5 — SI 53.9 · $6.00/1M blended · confidence 80% DeepSeek V3.2 — SI 53.3 · $0.31/1M blended · confidence 56% Step 3.5 Flash — SI 53.0 · $0.15/1M blended · confidence 88% Qwen3 30B A3B — SI 52.6 · $0.19/1M blended · confidence 64% o4-mini — SI 51.2 · $1.93/1M blended · confidence 87% GPT-5 Mini — SI 50.6 · $0.69/1M blended · confidence 99% Qwen3 32B — SI 50.4 · $0.28/1M blended · confidence 67% Qwen3 235B-A22B — SI 50.3 · $0.50/1M blended · confidence 76% Qwen3 Max — SI 50.2 · $0.63/1M blended · confidence 69% GPT-4o (2024-05-13) — SI 49.4 · $7.50/1M blended · confidence 93% DeepSeek-R1 — SI 47.9 · $0.96/1M blended · confidence 88% GLM-4.6 — SI 47.8 · $1.00/1M blended · confidence 55% Claude Sonnet 4 — SI 46.6 · $6.00/1M blended · confidence 100% GPT-5 Pro — SI 45.8 · $41.25/1M blended · confidence 64% Gemini 2.5 Pro — SI 45.2 · $3.44/1M blended · confidence 90% GPT-4o (2024-11-20) — SI 42.2 · $4.38/1M blended · confidence 61% o3-mini — SI 41.6 · $1.93/1M blended · confidence 96% GPT-4.1 — SI 41.5 · $3.50/1M blended · confidence 100% Claude Haiku 3.5 — SI 40.8 · $1.60/1M blended · confidence 100% GPT-5 Nano — SI 39.1 · $0.14/1M blended · confidence 85% GPT-4.1 mini — SI 36.9 · $0.70/1M blended · confidence 87% GPT-4o (2024-08-06) — SI 36.7 · $4.38/1M blended · confidence 100% GPT-4.1 nano — SI 32.1 · $0.17/1M blended · confidence 82% GPT-4o — SI 31.2 · $4.38/1M blended · confidence 59% GPT-4o mini — SI 22.9 · $0.26/1M blended · confidence 100%
Closed weights Open weights Dot opacity = confidence in the score
SI Score vs blended API price (log scale). Upper-left is the sweet spot: more score per dollar.
SI Score 20 40 60 80 100 Mar ’24 Oct ’24 May ’25 Dec ’25 Jul ’26 Release date Claude Fable 5.1 — SI 80.2 · released 2026-09-01 · confidence 100% Claude Opus 5.5 — SI 78.4 · released 2026-09-22 · confidence 100% GPT-6 Astra — SI 77.5 · released 2026-09-03 · confidence 100% Claude Fable 5 — SI 76.8 · released 2026-06-09 · confidence 100% Claude Opus 5 — SI 74.9 · released 2026-07-24 · confidence 100% GPT-6.1 Sol — SI 73.0 · released 2026-09-29 · confidence 100% GPT-5.6 Sol — SI 72.1 · released 2026-07-09 · confidence 100% Claude Opus 4.7 — SI 71.2 · released 2026-04-16 · confidence 100% Kimi K3 — SI 71.0 · released 2026-07-16 · confidence 100% Claude Opus 4.6 — SI 70.8 · released 2026-02-05 · confidence 100% Muse Spark 1.3 — SI 70.1 · released 2026-09-02 · confidence 93% Gemini 3.7 Flash — SI 70.0 · released 2026-08-13 · confidence 100% GPT-5.4 — SI 69.7 · released 2026-03-05 · confidence 100% GPT-5.5 — SI 69.5 · released 2026-04-24 · confidence 100% Claude Sonnet 5.5 — SI 69.4 · released 2026-09-28 · confidence 93% GPT-6 Sol — SI 68.9 · released 2026-09-22 · confidence 100% Gemini 3.8 Flash — SI 68.8 · released 2026-09-02 · confidence 100% Qwen3.8 Max — SI 68.7 · released 2026-08-03 · confidence 85% Gemini 3.5 Flash — SI 68.3 · released 2026-05-19 · confidence 100% Claude Sonnet 4.6 — SI 68.3 · released 2026-02-17 · confidence 100% GPT-5.6 Terra — SI 68.2 · released 2026-07-09 · confidence 100% Gemini 3.1 Pro Preview — SI 68.1 · released 2026-02-19 · confidence 100% MiniMax-M3 — SI 68.0 · released 2026-06-01 · confidence 100% GLM-5.3 — SI 67.7 · released 2026-08-14 · confidence 93% Claude Sonnet 5 — SI 67.7 · released 2026-06-30 · confidence 93% Claude Opus 4.8 — SI 67.5 · released 2026-05-28 · confidence 100% Grok 4.6 — SI 66.6 · released 2026-08-12 · confidence 100% Inkling — SI 66.5 · released 2026-07-15 · confidence 100% Qwen3.6 Max Preview — SI 66.4 · released 2026-04-20 · confidence 64% GLM-5.3-Flash — SI 66.2 · released 2026-08-26 · confidence 100% DeepSeek V4.1 Flash — SI 66.2 · released 2026-09-10 · confidence 69% Kimi K2 Thinking Turbo — SI 66.1 · released 2025-11-06 · confidence 64% GPT-5.2 — SI 65.9 · released 2025-12-11 · confidence 100% GLM-5.2 — SI 65.7 · released 2026-06-13 · confidence 100% Gemini 3.6 Flash — SI 65.5 · released 2026-07-21 · confidence 100% Grok 4.7 — SI 65.5 · released 2026-09-21 · confidence 100% Grok 4.5 — SI 64.8 · released 2026-07-08 · confidence 100% Kimi K2.6 — SI 64.0 · released 2026-04-21 · confidence 85% Qwen3.5 397B-A17B — SI 64.0 · released 2026-02-15 · confidence 69% GPT-5.5 Pro — SI 63.5 · released 2026-04-24 · confidence 53% GPT-5.6 Luna — SI 63.4 · released 2026-07-09 · confidence 100% Gemma 4 26B A4B IT — SI 63.4 · released 2026-04-02 · confidence 64% GLM-5.1 — SI 63.3 · released 2026-04-07 · confidence 64% DeepSeek V4 Pro — SI 63.1 · released 2026-04-24 · confidence 85% Gemma 4 31B IT — SI 63.0 · released 2026-04-02 · confidence 64% DeepSeek V4 Pro 0813 — SI 63.0 · released 2026-08-12 · confidence 69% Muse Spark 1.1 — SI 62.9 · released 2026-04-08 · confidence 61% Qwen3.7 Plus — SI 62.7 · released 2026-06-02 · confidence 64% Qwen3.5 35B-A3B — SI 62.7 · released 2026-02-23 · confidence 64% Claude Opus 4.5 — SI 62.6 · released 2025-11-01 · confidence 100% MiMo-V2.5-Pro — SI 62.6 · released 2026-04-22 · confidence 88% Muse Spark 1.2 — SI 62.3 · released 2026-08-05 · confidence 53% Mistral Large 4 — SI 62.2 · released 2026-10-06 · confidence 64% Qwen3.8 27B — SI 61.9 · released 2026-08-14 · confidence 69% GPT-6 Luna — SI 61.6 · released 2026-09-22 · confidence 100% Qwen3.6 Plus — SI 61.6 · released 2026-04-02 · confidence 80% GLM-4.7 — SI 61.3 · released 2025-12-22 · confidence 69% GPT-5.4 Pro — SI 61.3 · released 2026-03-05 · confidence 69% Nemotron 3 Ultra 550B A55B — SI 61.2 · released 2026-06-04 · confidence 100% DeepSeek V4 Flash 0731 — SI 61.0 · released 2026-07-31 · confidence 69% Hy3 — SI 60.7 · released 2026-07-06 · confidence 88% Claude Haiku 4.5 — SI 60.5 · released 2025-10-15 · confidence 64% Qwen3.5 Flash — SI 59.5 · released 2026-02-23 · confidence 64% Claude Haiku 5.5 — SI 59.4 · released 2026-10-07 · confidence 68% GLM-5 — SI 59.4 · released 2026-02-12 · confidence 90% Kimi K2.5 — SI 59.0 · released 2026-01 · confidence 100% Qwen3.7 Max — SI 58.6 · released 2026-05-21 · confidence 53% Gemini 3 Pro Preview — SI 58.6 · released 2025-11-18 · confidence 68% GPT-5.5 Instant — SI 58.3 · released 2026-05-05 · confidence 64% MiniMax-M2.7 — SI 57.9 · released 2026-03-18 · confidence 88% Fugu Ultra — SI 57.9 · released 2026-06-15 · confidence 100% MiniMax-M2.5 — SI 57.8 · released 2026-02-12 · confidence 100% Grok 4.20 (Reasoning) — SI 57.7 · released 2026-03-09 · confidence 80% MiMo-V2.6-Pro — SI 57.4 · released 2026-09-22 · confidence 88% GPT-5.4 nano — SI 57.4 · released 2026-03-17 · confidence 100% GPT-5.4 mini — SI 56.9 · released 2026-03-17 · confidence 100% LongCat-2.0 — SI 56.6 · released 2026-06-30 · confidence 100% MiMo-V2.5 — SI 56.4 · released 2026-04-22 · confidence 88% DeepSeek V4 Flash — SI 56.2 · released 2026-04-24 · confidence 53% Fugu — SI 55.9 · released 2026-06-15 · confidence 100% MiMo-V2.6-Flash — SI 55.8 · released 2026-09-22 · confidence 88% DeepSeek V3 0324 — SI 55.7 · released 2025-03-24 · confidence 74% GPT-5 — SI 55.7 · released 2025-08-07 · confidence 100% MiniMax-M2 — SI 55.7 · released 2025-10-27 · confidence 96% GPT-5.1 — SI 55.7 · released 2025-11-13 · confidence 85% Qwen3.6 27B — SI 55.6 · released 2026-04-22 · confidence 53% Claude Opus 4 — SI 55.5 · released 2025-05-22 · confidence 100% o3 — SI 55.4 · released 2025-04-16 · confidence 87% Gemini 3.5 Flash Lite — SI 55.2 · released 2026-07-21 · confidence 100% GPT OSS 120B — SI 54.8 · released 2025-08-05 · confidence 79% Gemini 3 Flash Preview — SI 54.8 · released 2025-12-17 · confidence 93% Grok 4.3 — SI 54.6 · released 2026-04-17 · confidence 80% DeepSeek-V3 — SI 54.1 · released 2024-12-26 · confidence 93% Claude Sonnet 4.5 — SI 53.9 · released 2025-09-29 · confidence 80% Inkling Small — SI 53.6 · released 2026-07-30 · confidence 100% DeepSeek V3.2 — SI 53.3 · released 2025-12-01 · confidence 56% GLM-4.7-Flash — SI 53.3 · released 2026-01-19 · confidence 69% Step 3.5 Flash — SI 53.0 · released 2026-01-29 · confidence 88% Qwen3 30B A3B — SI 52.6 · released 2025-04-28 · confidence 64% GPT OSS 20B — SI 52.5 · released 2025-08-05 · confidence 64% Claude Sonnet 3.5 v2 — SI 51.9 · released 2024-10-22 · confidence 100% Claude Opus 4.1 — SI 51.5 · released 2025-08-05 · confidence 80% o4-mini — SI 51.2 · released 2025-04-16 · confidence 87% Llama 4 Scout 17B Instruct — SI 51.0 · released 2025-04-05 · confidence 64% GPT-5 Mini — SI 50.6 · released 2025-08-07 · confidence 99% Qwen3 32B — SI 50.4 · released 2025-04 · confidence 67% Qwen3 235B-A22B — SI 50.3 · released 2025-04 · confidence 76% Qwen3 Max — SI 50.2 · released 2025-09-23 · confidence 69% GPT-4o (2024-05-13) — SI 49.4 · released 2024-05-13 · confidence 93% Claude Sonnet 3.7 — SI 48.9 · released 2025-02-19 · confidence 100% Mistral Small 3.1 24B — SI 48.4 · released 2025-03-17 · confidence 64% QwQ 32B — SI 48.2 · released 2025-03-05 · confidence 67% DeepSeek-R1 — SI 47.9 · released 2025-01-20 · confidence 88% GLM-4.6 — SI 47.8 · released 2025-09-30 · confidence 55% Claude Sonnet 4 — SI 46.6 · released 2025-05-22 · confidence 100% Nemotron 3.5 Lightning 30B A3B — SI 46.1 · released 2026-08-11 · confidence 100% GPT-5 Pro — SI 45.8 · released 2025-10-06 · confidence 64% Llama-3.1-70B-Instruct — SI 45.7 · released 2024-07-23 · confidence 93% Gemini 2.5 Pro — SI 45.2 · released 2025-06-05 · confidence 90% Gemma 3 12B IT — SI 45.1 · released 2025-03-12 · confidence 64% GPT-4o (2024-11-20) — SI 42.2 · released 2024-11-20 · confidence 61% Gemma 3 27B IT — SI 42.0 · released 2025-03-12 · confidence 74% o3-mini — SI 41.6 · released 2024-12-20 · confidence 96% GPT-4.1 — SI 41.5 · released 2025-04-14 · confidence 100% Claude Haiku 3.5 — SI 40.8 · released 2024-10-22 · confidence 100% Claude Haiku 3 — SI 39.9 · released 2024-03-13 · confidence 93% Qwen2.5-Coder-32B-Instruct — SI 39.6 · released 2024-11-12 · confidence 90% Gemma 3 4B IT — SI 39.2 · released 2025-03-12 · confidence 64% GPT-5 Nano — SI 39.1 · released 2025-08-07 · confidence 85% Llama-3.1-8B-Instruct — SI 37.7 · released 2024-07-23 · confidence 93% Nova Pro — SI 37.4 · released 2024-12-03 · confidence 100% GPT-4.1 mini — SI 36.9 · released 2025-04-14 · confidence 87% GPT-4o (2024-08-06) — SI 36.7 · released 2024-08-06 · confidence 100% Nova Lite — SI 36.6 · released 2024-12-03 · confidence 100% Mistral Large 2.1 — SI 35.3 · released 2024-11-18 · confidence 100% Mistral Medium 3 — SI 33.0 · released 2025-05-07 · confidence 85% Llama-3.3-70B-Instruct — SI 32.7 · released 2024-12-06 · confidence 100% Llama-3.2-1B — SI 32.1 · released 2024-09-25 · confidence 93% GPT-4.1 nano — SI 32.1 · released 2025-04-14 · confidence 82% GPT-4o — SI 31.2 · released 2024-05-13 · confidence 59% Llama 4 Maverick 17B Instruct — SI 30.0 · released 2025-04-05 · confidence 78% GPT-4o mini — SI 22.9 · released 2024-07-18 · confidence 100%
Closed weights Open weights Dot opacity = confidence in the score
Ranked models by release date and current SI Score. This is a release comparison, not historical score tracking.

Latest releases

All releases →

Understand the numbers