SI Score · updated Oct 9, 2026
Superintelligence leaderboard.
Open evidence. Clear rankings.
The latest frontier models, ranked by the SI Score — a composite of open benchmarks, scaled per benchmark and weighted across reasoning, coding, math and human preference. Every number links to its source and date, and each score carries a confidence % that rises as sources report.
The leaderboard
Full catalog →| 1 | Claude Fable 5.1 Anthropic | 80.2 | 100% confidence 100 percent, Full | $10.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | — |
| 2 | Claude Opus 5.5 Anthropic | 78.4 | 100% confidence 100 percent, Full | $4.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | — |
| 3 | GPT-6 Astra OpenAI | 77.5 | 100% confidence 100 percent, Full | $10.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 4 | Claude Fable 5 Anthropic | 76.8 | 100% confidence 100 percent, Full | $10.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 5 | Claude Opus 5 Anthropic | 74.9 | 100% confidence 100 percent, Full | $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 6 | GPT-6.1 Sol OpenAI | 73.0 | 100% confidence 100 percent, Full | $2.00OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 7 | GPT-5.6 Sol OpenAI | 72.1 | 100% confidence 100 percent, Full | $4.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-sol
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 8 | Claude Opus 4.7 Anthropic | 71.2 | 100% confidence 100 percent, Full | $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 9 | Kimi K3 Moonshot AI | 71.0 | 100% confidence 100 percent, Full | $3.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k3
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 10 | Claude Opus 4.6 Anthropic | 70.8 | 100% confidence 100 percent, Full | $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 11 | Muse Spark 1.3 Meta | 70.1 | 93% confidence 93 percent, High | $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 12 | Gemini 3.7 Flash Google | 70.0 | 100% confidence 100 percent, Full | $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 13 | GPT-5.4 OpenAI | 69.7 | 100% confidence 100 percent, Full | $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 14 | GPT-5.5 OpenAI | 69.5 | 100% confidence 100 percent, Full | $5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 15 | Claude Sonnet 5.5 Anthropic | 69.4 | 93% confidence 93 percent, High | $2.00Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | — |
| 16 | GPT-6 Sol OpenAI | 68.9 | 100% confidence 100 percent, Full | $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-6-sol
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 17 | Gemini 3.8 Flash Google | 68.8 | 100% confidence 100 percent, Full | $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 18 | Qwen3.8 Max Alibaba / Qwen | 68.7 | 85% confidence 85 percent, High | $1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 19 | Gemini 3.5 Flash Google | 68.3 | 100% confidence 100 percent, Full | $1.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 20 | Claude Sonnet 4.6 Anthropic | 68.3 | 100% confidence 100 percent, Full | $3.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 21 | GPT-5.6 Terra OpenAI | 68.2 | 100% confidence 100 percent, Full | $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-terra
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 22 | Gemini 3.1 Pro Preview Google | 68.1 | 100% confidence 100 percent, Full | $2.00Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 23 | MiniMax-M3 minimax | 68.0 | 100% confidence 100 percent, Full | $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 24 | GLM-5.3 Z.ai | 67.7 | 93% confidence 93 percent, High | $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 25 | Claude Sonnet 5 Anthropic | 67.7 | 93% confidence 93 percent, High | $2.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 26 | Claude Opus 4.8 Anthropic | 67.5 | 100% confidence 100 percent, Full | $5.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 27 | Grok 4.6 xAI | 66.6 | 100% confidence 100 percent, Full | $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.6
Retrieved Oct 9, 2026 · MIT Open source ↗ | 500Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 28 | Inkling thinkingmachines | 66.5 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 29 | Qwen3.6 Max Preview Alibaba / Qwen | 66.4 | 64% confidence 64 percent, Medium | $1.30Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤128K; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 30 | GLM-5.3-Flash Z.ai | 66.2 | 100% confidence 100 percent, Full | $0.15Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 31 | DeepSeek V4.1 Flash DeepSeek | 66.2 | 69% confidence 69 percent, Medium | $0.15DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 32 | Kimi K2 Thinking Turbo Moonshot AI | 66.1 | 64% confidence 64 percent, Medium | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 33 | GPT-5.2 OpenAI | 65.9 | 100% confidence 100 percent, Full | $1.75models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.2
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 34 | GLM-5.2 Z.ai | 65.7 | 100% confidence 100 percent, Full | $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 35 | Gemini 3.6 Flash Google | 65.5 | 100% confidence 100 percent, Full | $0.75Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 36 | Grok 4.7 xAI | 65.5 | 100% confidence 100 percent, Full | $2.00xAI models & pricingPublished source fact
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 500KxAI models & pricingPublished source fact
Retrieved Oct 9, 2026 · factual citation Open source ↗ | — |
| 37 | Grok 4.5 xAI | 64.8 | 100% confidence 100 percent, Full | $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.5
Retrieved Oct 9, 2026 · MIT Open source ↗ | 500Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 38 | Kimi K2.6 Moonshot AI | 64.0 | 85% confidence 85 percent, High | $0.95models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.moonshot.ai/docs/api/chat. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; moonshotai/kimi-k2.6
Retrieved Oct 9, 2026 · MIT Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 39 | Qwen3.5 397B-A17B Alibaba / Qwen | 64.0 | 69% confidence 69 percent, Medium | $0.17Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 40 | GPT-5.5 Pro OpenAI | 63.5 | 53% confidence 53 percent, Medium | $30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.5-pro
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 41 | GPT-5.6 Luna OpenAI | 63.4 | 100% confidence 100 percent, Full | $0.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.6-luna
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 42 | Gemma 4 26B A4B IT Google | 63.4 | 64% confidence 64 percent, Medium | $0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 43 | GLM-5.1 Z.ai | 63.3 | 64% confidence 64 percent, Medium | $1.40Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 44 | DeepSeek V4 Pro DeepSeek | 63.1 | 85% confidence 85 percent, High | $1.32LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 45 | Gemma 4 31B IT Google | 63.0 | 64% confidence 64 percent, Medium | $0.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://ai.google.dev/gemini-api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 46 | DeepSeek V4 Pro 0813 DeepSeek | 63.0 | 69% confidence 69 percent, Medium | $0.66DeepSeek pricingOfficial off-peak uncached rate; peak is 2x; time schedule at source
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 47 | Muse Spark 1.1 Meta | 62.9 | 61% confidence 61 percent, Medium | $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 48 | Qwen3.7 Plus Alibaba / Qwen | 62.7 | 64% confidence 64 percent, Medium | $0.28Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 49 | Qwen3.5 35B-A3B Alibaba / Qwen | 62.7 | 64% confidence 64 percent, Medium | $0.06Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 50 | Claude Opus 4.5 Anthropic | 62.6 | 100% confidence 100 percent, Full | $5.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-opus-4-5-20251101
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 51 | MiMo-V2.5-Pro xiaomi | 62.6 | 88% confidence 88 percent, High | $0.43models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5-pro
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 52 | Muse Spark 1.2 Meta | 62.3 | 53% confidence 53 percent, Medium | $1.25Meta API pricingOfficial Meta Standard tier; applies only to the versions explicitly listed by the provider. Contributor training-data-discount tier and cached input excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 53 | Mistral Large 4 mistral | 62.2 | 64% confidence 64 percent, Medium | $0.68Mistral API pricingOfficial Mistral Serverless API Standard rate, displayed sale price where applicable; cache, Batch, specialist units and hosted third-party models excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 54 | Qwen3.8 27B Alibaba / Qwen | 61.9 | 69% confidence 69 percent, Medium | $0.50Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 55 | GPT-6 Luna OpenAI | 61.6 | 100% confidence 100 percent, Full | $0.10OpenAI pricingOfficial Standard short-context rate; excludes Batch/Flex/cache discounts
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 56 | Qwen3.6 Plus Alibaba / Qwen | 61.6 | 80% confidence 80 percent, High | $0.28Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 57 | GLM-4.7 Z.ai | 61.3 | 69% confidence 69 percent, Medium | $0.60Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 205Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 58 | GPT-5.4 Pro OpenAI | 61.3 | 69% confidence 69 percent, Medium | $30.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-pro
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1.1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 59 | Nemotron 3 Ultra 550B A55B nvidia | 61.2 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 60 | DeepSeek V4 Flash 0731 DeepSeek | 61.0 | 69% confidence 69 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 61 | Hy3 tencent | 60.7 | 88% confidence 88 percent, High | $0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://cloud.tencent.com/document/product/1823/130050. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; tencent-tokenhub/hy3
Retrieved Oct 9, 2026 · MIT Open source ↗ | 256Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 62 | Claude Haiku 4.5 Anthropic | 60.5 | 64% confidence 64 percent, Medium | $1.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-haiku-4-5-20251001
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 63 | Qwen3.5 Flash Alibaba / Qwen | 59.5 | 64% confidence 64 percent, Medium | $0.03Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤128K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 64 | Claude Haiku 5.5 Anthropic | 59.4 | 68% confidence 68 percent, Medium | $0.10Anthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1MAnthropic models & pricingOfficial Claude API; lowest short-context on-demand tier; batch/cache/long-context rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | — |
| 65 | GLM-5 Z.ai | 59.4 | 90% confidence 90 percent, High | $1.00Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 205Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 66 | Kimi K2.5 Moonshot AI | 59.0 | 100% confidence 100 percent, Full | $0.60LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://platform.moonshot.ai/docs/guide/kimi-k2-5-quickstart. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 67 | Qwen3.7 Max Alibaba / Qwen | 58.6 | 53% confidence 53 percent, Medium | $1.65Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤1M; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 68 | Gemini 3 Pro Preview Google | 58.6 | 68% confidence 68 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 69 | GPT-5.5 Instant OpenAI | 58.3 | 64% confidence 64 percent, Medium | not yet reported | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 70 | MiniMax-M2.7 minimax | 57.9 | 88% confidence 88 percent, High | $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 205Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 71 | Fugu Ultra sakana | 57.9 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 72 | MiniMax-M2.5 minimax | 57.8 | 100% confidence 100 percent, Full | $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 205Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 73 | Grok 4.20 (Reasoning) xAI | 57.7 | 80% confidence 80 percent, High | $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.20-0309-reasoning
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 74 | MiMo-V2.6-Pro xiaomi | 57.4 | 88% confidence 88 percent, High | $0.43models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-pro
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 75 | GPT-5.4 nano OpenAI | 57.4 | 100% confidence 100 percent, Full | $0.20models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-nano
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 76 | GPT-5.4 mini OpenAI | 56.9 | 100% confidence 100 percent, Full | $0.75models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.4-mini
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 77 | LongCat-2.0 meituan | 56.6 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 78 | MiMo-V2.5 xiaomi | 56.4 | 88% confidence 88 percent, High | $0.14models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.5
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 79 | DeepSeek V4 Flash DeepSeek | 56.2 | 53% confidence 53 percent, Medium | $0.30LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://api-docs.deepseek.com/quick_start/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 80 | Fugu sakana | 55.9 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 81 | MiMo-V2.6-Flash xiaomi | 55.8 | 88% confidence 88 percent, High | $0.14models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.xiaomimimo.com/#/docs. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xiaomi/mimo-v2.6-flash
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 82 | DeepSeek V3 0324 DeepSeek | 55.7 | 74% confidence 74 percent, Medium | not yet reported | 164Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 83 | GPT-5 OpenAI | 55.7 | 100% confidence 100 percent, Full | $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 84 | MiniMax-M2 minimax | 55.7 | 96% confidence 96 percent, High | $0.30MiniMax API pricingOfficial MiniMax global on-demand API; lowest short-context tier and displayed permanent promotional discount; excludes high-context, fast tier and subscriptions
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 205Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 85 | GPT-5.1 OpenAI | 55.7 | 85% confidence 85 percent, High | $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5.1
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 86 | Qwen3.6 27B Alibaba / Qwen | 55.6 | 53% confidence 53 percent, Medium | $0.60Alibaba Model Studio pricingOfficial Alibaba Model Studio International USD on-demand API; 0<Token≤256K; . Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 87 | Claude Opus 4 Anthropic | 55.5 | 100% confidence 100 percent, Full | $15.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 88 | o3 OpenAI | 55.4 | 87% confidence 87 percent, High | $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/o3
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 89 | Gemini 3.5 Flash Lite Google | 55.2 | 100% confidence 100 percent, Full | $0.30Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 90 | GPT OSS 120B OpenAI | 54.8 | 79% confidence 79 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 91 | Gemini 3 Flash Preview Google | 54.8 | 93% confidence 93 percent, High | $0.50Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 92 | Grok 4.3 xAI | 54.6 | 80% confidence 80 percent, High | $1.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.x.ai/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; xai/grok-4.3
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 93 | DeepSeek-V3 DeepSeek | 54.1 | 93% confidence 93 percent, High | $0.27LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 94 | Claude Sonnet 4.5 Anthropic | 53.9 | 80% confidence 80 percent, High | $3.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.anthropic.com/en/docs/about-claude/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; anthropic/claude-sonnet-4-5-20250929
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 95 | Inkling Small thinkingmachines | 53.6 | 100% confidence 100 percent, Full | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 96 | DeepSeek V3.2 DeepSeek | 53.3 | 56% confidence 56 percent, Medium | $0.28LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 97 | GLM-4.7-Flash Z.ai | 53.3 | 69% confidence 69 percent, Medium | $0.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://docs.z.ai/guides/overview/pricing. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; zai/glm-4.7-flash
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 98 | Step 3.5 Flash stepfun | 53.0 | 88% confidence 88 percent, High | $0.10models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.stepfun.com/docs/zh/overview/concept. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; stepfun/step-3.5-flash
Retrieved Oct 9, 2026 · MIT Open source ↗ | 256Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 99 | Qwen3 30B A3B Alibaba / Qwen | 52.6 | 64% confidence 64 percent, Medium | $0.11Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 100 | GPT OSS 20B OpenAI | 52.5 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 101 | Claude Sonnet 3.5 v2 Anthropic | 51.9 | 100% confidence 100 percent, Full | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 102 | Claude Opus 4.1 Anthropic | 51.5 | 80% confidence 80 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 103 | o4-mini OpenAI | 51.2 | 87% confidence 87 percent, High | $1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 104 | Llama 4 Scout 17B Instruct Meta | 51.0 | 64% confidence 64 percent, Medium | not yet reported | 10Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 105 | GPT-5 Mini OpenAI | 50.6 | 99% confidence 99 percent, High | $0.25models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-mini
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 106 | Qwen3 32B Alibaba / Qwen | 50.4 | 67% confidence 67 percent, Medium | $0.16Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 107 | Qwen3 235B-A22B Alibaba / Qwen | 50.3 | 76% confidence 76 percent, Medium | $0.29Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; standard tier; Non-Thinking and Thinking modes. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 108 | Qwen3 Max Alibaba / Qwen | 50.2 | 69% confidence 69 percent, Medium | $0.36Alibaba Model Studio pricingOfficial Alibaba Model Studio Global USD on-demand API; 0<Token≤32K; Non-Thinking mode only. Output uses non-thinking rate when both modes are available; thinking-only products use their thinking rate. Cache, Batch, free quotas and regional rates excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 109 | GPT-4o (2024-05-13) OpenAI | 49.4 | 93% confidence 93 percent, High | $5.00LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 110 | Claude Sonnet 3.7 Anthropic | 48.9 | 100% confidence 100 percent, Full | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 111 | Mistral Small 3.1 24B mistral | 48.4 | 64% confidence 64 percent, Medium | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 112 | QwQ 32B Alibaba / Qwen | 48.2 | 67% confidence 67 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 113 | DeepSeek-R1 DeepSeek | 47.9 | 88% confidence 88 percent, High | $0.55LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: not supplied in the MIT entry; rate is a transcription, not independently verified. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 114 | GLM-4.6 Z.ai | 47.8 | 55% confidence 55 percent, Medium | $0.60Z.ai API pricingOfficial Z.ai global Standard uncached token rate; coding-plan subscriptions, cache and Batch excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 205Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 115 | Claude Sonnet 4 Anthropic | 46.6 | 100% confidence 100 percent, Full | $3.00Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 116 | Nemotron 3.5 Lightning 30B A3B nvidia | 46.1 | 100% confidence 100 percent, Full | not yet reported | 262Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 117 | GPT-5 Pro OpenAI | 45.8 | 64% confidence 64 percent, Medium | $15.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-pro
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 118 | Llama-3.1-70B-Instruct Meta | 45.7 | 93% confidence 93 percent, High | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 119 | Gemini 2.5 Pro Google | 45.2 | 90% confidence 90 percent, High | $1.25Google Gemini pricingOfficial paid Standard text rate, lowest short-context tier; current promotional price if dated; excludes free/Batch/Flex/audio
Retrieved Oct 9, 2026 · CC-BY-4.0 factual citation Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 120 | Gemma 3 12B IT Google | 45.1 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 121 | GPT-4o (2024-11-20) OpenAI | 42.2 | 61% confidence 61 percent, Medium | $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-11-20
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 122 | Gemma 3 27B IT Google | 42.0 | 74% confidence 74 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 123 | o3-mini OpenAI | 41.6 | 96% confidence 96 percent, High | $1.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 124 | GPT-4.1 OpenAI | 41.5 | 100% confidence 100 percent, Full | $2.00models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 125 | Claude Haiku 3.5 Anthropic | 40.8 | 100% confidence 100 percent, Full | $0.80Anthropic API pricingOfficial Claude API base tokens; lowest short-context global Standard rate; cache, batch, fast and regional premiums excluded
Retrieved Oct 9, 2026 · factual citation Open source ↗ | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 126 | Claude Haiku 3 Anthropic | 39.9 | 93% confidence 93 percent, High | not yet reported | 200Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 127 | Qwen2.5-Coder-32B-Instruct Alibaba / Qwen | 39.6 | 90% confidence 90 percent, High | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 128 | Gemma 3 4B IT Google | 39.2 | 64% confidence 64 percent, Medium | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 129 | GPT-5 Nano OpenAI | 39.1 | 85% confidence 85 percent, High | $0.05models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-5-nano
Retrieved Oct 9, 2026 · MIT Open source ↗ | 400Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 130 | Llama-3.1-8B-Instruct Meta | 37.7 | 93% confidence 93 percent, High | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 131 | Nova Pro amazon | 37.4 | 100% confidence 100 percent, Full | not yet reported | 300Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 132 | GPT-4.1 mini OpenAI | 36.9 | 87% confidence 87 percent, High | $0.40models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4.1-mini
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 133 | GPT-4o (2024-08-06) OpenAI | 36.7 | 100% confidence 100 percent, Full | $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-2024-08-06
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 134 | Nova Lite amazon | 36.6 | 100% confidence 100 percent, Full | not yet reported | 300Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 135 | Mistral Large 2.1 mistral | 35.3 | 100% confidence 100 percent, Full | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 136 | Mistral Medium 3 mistral | 33.0 | 85% confidence 85 percent, High | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 137 | Llama-3.3-70B-Instruct Meta | 32.7 | 100% confidence 100 percent, Full | not yet reported | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 138 | Llama-3.2-1B Meta | 32.1 | 93% confidence 93 percent, High | not yet reported | 131Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 139 | GPT-4.1 nano OpenAI | 32.1 | 82% confidence 82 percent, High | $0.10LiteLLMFirst-party API Standard token rate; MIT LiteLLM transcription. Provider documentation: https://developers.openai.com/api/docs/pricing. Exact endpoint only; cache/batch/long-context rates excluded
Retrieved Oct 9, 2026 · MIT Open source ↗ | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 140 | GPT-4o OpenAI | 31.2 | 59% confidence 59 percent, Medium | $2.50models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
| 141 | Llama 4 Maverick 17B Instruct Meta | 30.0 | 78% confidence 78 percent, Medium | not yet reported | 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | Open |
| 142 | GPT-4o mini OpenAI | 22.9 | 100% confidence 100 percent, Full | $0.15models.devFirst-party hosted API; MIT models.dev transcription. Provider documentation: https://platform.openai.com/docs/models. Exact canonical endpoint; lowest short-context Standard USD token tier; cache/batch discounts excluded. Deprecated endpoints excluded; openai/gpt-4o-mini
Retrieved Oct 9, 2026 · MIT Open source ↗ | 128Kmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT Open source ↗ | — |
Ranked text models appear first. The provisional toggle includes only text models awaiting evidence. Browse every kind in the model catalog. Prices are USD per 1M tokens (official first-party API). Dotted values carry their source — hover or tap to see it. “not yet reported” means the source has not published that figure for this model.
Score against price, and over time
Best value models →Latest releases
All releases →Oct 7, 2026
Claude Haiku 5.5
Anthropic
SI 59.4 · confidence 68%
Oct 6, 2026
Nano Banana 2.1
Google · provisional
score not yet reported
Oct 6, 2026
Mistral Large 4
mistral · provisional
SI 62.2 · confidence 64%
Oct 1, 2026
Grok Imagine Video 1.5 Lite
xAI · provisional
score not yet reported
Sep 29, 2026
Ling 3.1 Flash
inclusionai · provisional
score not yet reported
Understand the numbers
Methodology
How the SI Score is built, what the confidence % means, which sources we use and their licenses.
What is superintelligence?
The research term, the corporate lane-name, and the new government spelling — disentangled and dated.
“Super Intelligence” in U.S. policy
Executive Order 14434 made it executive-branch vocabulary on Sep 29, 2026. What it says, verbatim.
Statements tracker
Who has adopted the term, who rejects it, and what comes next — dated, sourced, neutral.