Pareto
Flagshipby Unbiased
Pareto is the model most people met under a different name. On September 16, 2026 a listing called Union Alpha appeared on OpenRouter with no lab attached, free, multimodal, 262K context, and it drew enough traffic in a day to degrade its own latency; capacity was tripled overnight and still ran short. At 23:24 UTC on September 17 the stealth listing lost its endpoints and the model went live under its real name as Pareto 26.9 from Unbiased, the platform built by Circuit & Chisel, at $2.50 per million input tokens, $0.25 cached, and $7.50 output. The architecture is the part to read carefully, because Pareto is not one model. Unbiased describes it as a blended model: several frontier and open-source models run against each request and the best answer is kept, behind one model string and one bill, and it never switches models mid-conversation so prompt caching keeps working. That design is a real hedge against any single vendor moving its price or revoking access, and it is also the reason the usual questions have no answer here, since Unbiased publishes no parameter count, no architecture, and no list of which models are in the blend. Specs are a 262,144 token context window with up to 131,072 output, text and image input, text output, and tool calling with structured output. On its own published table Pareto scores 88 on ArXivMath, 78 on MMMU-Pro, and ties GPT-6 Astra and DeepSeek V4.1 Flash at 74 on DeepSWE. No independent evaluator has published a score, and Artificial Analysis has no entry, so treat the whole table as vendor-reported. Available on the Unbiased platform as pareto, and through OpenRouter, Cloudflare AI Gateway, Kilo Gateway, and NanoGPT at the same rate.
Input Price
$2.50
per 1M tokens
Output Price
$7.50
per 1M tokens
Context Window
262K
tokens
Released
2026-09
API access
Capabilities
Key Strengths
- ✓Blends several frontier and open models per request behind one model string
- ✓Never switches models mid-conversation, so prompt caching still applies
- ✓$2.50/$7.50 with cached input at $0.25
- ✓262,144 token context with 131,072 max output
- ✓Tool calling with structured output
- ✓Reached production demand as the anonymous Union Alpha before it was named
Best For
- ▸Teams that want frontier-grade output without committing to one lab
- ▸Research and coding workloads with mixed task types
- ▸Agentic work where a single vendor outage is an unacceptable risk
- ▸Evaluation against single-model flagships at the same price point
Pricing Details
Input tokens
$2.50
per 1M tokens
Output tokens
$7.50
per 1M tokens
Estimated cost per 1K requests
$6.25
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.