Fugu Ultra v2.0
Flagshipby Sakana AI
Fugu Ultra v2.0 shipped on September 11, 2026, and the first thing to get straight is that it is not a model. It is an orchestrator: one request to one API, and Fugu decides which models in its pool do the work and stitches the results back together. Pricing is $5 per million input tokens, $30 output, and $0.50 cached input, stepping up to $10, $45, and $1.00 once context passes 272K tokens, with a 1 million token context window, and it speaks OpenAI-compatible Chat Completions and Responses plus an Anthropic-compatible Messages endpoint, so switching costs are a parameter change rather than a migration. Sakana reports best or joint-best on five of eight benchmarks (GDP.pdf, Chartography, DeepSWE, Toolathon, and the internal SWEFish) and top two on seven of eight, with Chartography at 48.3 against Opus 5 at 27.3 and Fable 5 at 29.5, and DeepSWE at 74.3. The footnote on Sakana's own chart is the interesting part: Fable 5, Fable 5.1, and GPT-6 Astra are not in the pool, and the training cutoff is August 28, 2026, so these numbers come from orchestrating open and specialized models rather than from reselling frontier access. That is also the pitch, since a swappable pool is insulation against vendor lock-in, price moves, and sudden API revocation. Read the billing model before you commit: orchestration tokens are reported separately inside token_details but bill at full input and output rates, so the invoice covers reasoning work the user never sees, and a per-request cost cannot be derived from visible tokens alone. Fugu is not yet offered in the EU or EEA. All benchmark figures are self-reported and SWEFish is Sakana's own.
Input Price
$5.00
per 1M tokens
Output Price
$30.00
per 1M tokens
Context Window
1M
tokens
Released
2026-09
API access
Capabilities
Key Strengths
- ✓Orchestrates a pool of models rather than serving one
- ✓Chartography 48.3 against Opus 5 at 27.3 and Fable 5 at 29.5
- ✓Scores without Fable 5, Fable 5.1, or GPT-6 Astra in its pool
- ✓1M token context window
- ✓OpenAI and Anthropic compatible endpoints
- ✓Swappable pool reduces single-vendor exposure
Best For
- ▸Complex multi-step reasoning and autonomous research
- ▸Full-stack software development agents
- ▸Visual and structured data interpretation
- ▸Teams that want frontier output without frontier vendor lock-in
Benchmark Scores
| Benchmark | Score | Description |
|---|---|---|
| GPQA Diamond | 95.5 | Graduate-level science questions verified by domain experts |
Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.
Pricing Details
Input tokens
$5.00
per 1M tokens
Output tokens
$30.00
per 1M tokens
Estimated cost per 1K requests
$20.00
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.