Skip to content
All systems operational0 AI providers monitored, polled every 2 minutes
Live status

Fugu Ultra v2.0

Flagship

by Sakana AI

Fugu Ultra v2.0 shipped on September 11, 2026, and the first thing to get straight is that it is not a model. It is an orchestrator: one request to one API, and Fugu decides which models in its pool do the work and stitches the results back together. Pricing is $5 per million input tokens, $30 output, and $0.50 cached input, stepping up to $10, $45, and $1.00 once context passes 272K tokens, with a 1 million token context window, and it speaks OpenAI-compatible Chat Completions and Responses plus an Anthropic-compatible Messages endpoint, so switching costs are a parameter change rather than a migration. Sakana reports best or joint-best on five of eight benchmarks (GDP.pdf, Chartography, DeepSWE, Toolathon, and the internal SWEFish) and top two on seven of eight, with Chartography at 48.3 against Opus 5 at 27.3 and Fable 5 at 29.5, and DeepSWE at 74.3. The footnote on Sakana's own chart is the interesting part: Fable 5, Fable 5.1, and GPT-6 Astra are not in the pool, and the training cutoff is August 28, 2026, so these numbers come from orchestrating open and specialized models rather than from reselling frontier access. That is also the pitch, since a swappable pool is insulation against vendor lock-in, price moves, and sudden API revocation. Read the billing model before you commit: orchestration tokens are reported separately inside token_details but bill at full input and output rates, so the invoice covers reasoning work the user never sees, and a per-request cost cannot be derived from visible tokens alone. Fugu is not yet offered in the EU or EEA. All benchmark figures are self-reported and SWEFish is Sakana's own.

Input Price

$5.00

per 1M tokens

Output Price

$30.00

per 1M tokens

Context Window

1M

tokens

Released

2026-09

API access

Capabilities

textvisiontool-usecodereasoning

Key Strengths

  • Orchestrates a pool of models rather than serving one
  • Chartography 48.3 against Opus 5 at 27.3 and Fable 5 at 29.5
  • Scores without Fable 5, Fable 5.1, or GPT-6 Astra in its pool
  • 1M token context window
  • OpenAI and Anthropic compatible endpoints
  • Swappable pool reduces single-vendor exposure

Best For

  • Complex multi-step reasoning and autonomous research
  • Full-stack software development agents
  • Visual and structured data interpretation
  • Teams that want frontier output without frontier vendor lock-in

Benchmark Scores

BenchmarkScoreDescription
GPQA Diamond95.5Graduate-level science questions verified by domain experts

Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.

Pricing Details

Input tokens

$5.00

per 1M tokens

Output tokens

$30.00

per 1M tokens

Estimated cost per 1K requests

$20.00

~1K input + ~500 output tokens avg

Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.

Related Models

View DocumentationCompare ModelsCost CalculatorFull Pricing Guide