Sakana AI, released Jun 15, 2026
Fugu Ultraprice, context, benchmarks and release details
- Input, per 1M tokens
- not yet reported
- Output, per 1M tokens
- not yet reported
- Context window
- 1Mmodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT
Open source ↗ - Max output
- not yet reported
- Released
- Jun 15, 2026models.devPublished source fact
Retrieved Oct 9, 2026 · MIT
Open source ↗
How the score breaks down
Weights: reasoning 30%, math 15%, coding 40%, preference 15%. Results use fixed 0–100 scales before averaging, and thin evidence is pulled toward 50. Method si-v3-retained-evidence-2, computed Oct 9, 2026, 06:15 UTC.
Around it on the leaderboard
- 69 GPT-5.5 Instant 58.3
- 70 MiniMax-M2.7 57.9
- 71 Fugu Ultra 57.9
- 72 MiniMax-M2.5 57.8
- 73 Grok 4.20 (Reasoning) 57.7
Benchmark results
4 benchmarks, 4 resultsEach row shows the best published result. Where a model was tested at several settings, such as reasoning effort, open the row to see each one. Hover or tap a value for its source.
Open source ↗ 95.5
Open source ↗ 50.0
Open source ↗ 73.7
Open source ↗ 82.1
Normalization uses a fixed 0–100 scale for each unit, independent of other models. Compare evaluation conditions before reading a small gap as decisive. “Lab-reported” marks the provider's own published figure.
Details and sources
- Open weights
- Nomodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT
Open source ↗ - License
- not yet reported
- Input modalities
- text, imagemodels.devPublished source fact
Retrieved Oct 9, 2026 · MIT
Open source ↗ - First seen by SuperIndex
- Oct 8, 2026
- Coverage
- 100% of expected source weight
Reported (1)
- Official model cards via models.devOct 8, 2026
Awaiting (0)
Every expected source has reported for this model.
Confidence rises as pending sources publish. Some sources never cover some models, so confidence reaches 100% at 80% of expected weight.