MAI-Code-1.1-Flash
Budgetby Microsoft
MAI-Code-1.1-Flash is Microsoft's second in-house coding model, released August 11, 2026, ten weeks after MAI-Code-1-Flash debuted at Build, and the headline is price. GitHub lists it at $0.20 per million input tokens, $0.02 cached, and $1.20 output, exactly a quarter of the $0.75, $0.075, and $4.50 that 1.0 charged. The model card describes a sparse mixture-of-experts with 138 billion total parameters and 5 billion active, a 256K token context window, and text plus image input, which 1.0 did not take. It started from the compressed 5B-active mid-training checkpoint of MAI-Thinking-1 rather than a fresh pretrain. Microsoft reports a 22 percent improvement on Terminal-Bench 2.1 inside GitHub Copilot CLI, a 15 percent gain on .NET tasks, 25 percent fewer tokens per completed task, 25 percent faster streaming, and code survival up 4 percent in production, all measured on its own harness and none independently reproduced. Weights are closed, and the license is whatever product terms apply where it is deployed, which in practice means GitHub Copilot first. GitHub deprecated MAI-Code-1-Flash on September 10, 2026, so this is now the Microsoft coding model to budget against Claude Haiku 4.5 and the Gemini Flash tier.
Input Price
$0.20
per 1M tokens
Output Price
$1.20
per 1M tokens
Context Window
256K
tokens
Released
2026-08
API access
Capabilities
Key Strengths
- ✓$0.20/$1.20 per 1M tokens, a quarter of MAI-Code-1-Flash
- ✓138B total with only 5B active parameters
- ✓256K context with new image input
- ✓25 percent fewer tokens per task in Copilot, Microsoft-reported
- ✓The Microsoft coding model after 1.0 was deprecated on September 10
Best For
- ▸High-volume Copilot coding and refactoring
- ▸CLI and terminal-heavy agent loops
- ▸.NET-heavy codebases
- ▸Screenshot and UI-mockup driven front-end work
Benchmark Scores
| Benchmark | Score | Description |
|---|---|---|
| SWE-bench | 72.6 | Real-world software engineering tasks from GitHub issues (SWE-bench Verified) |
Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.
Pricing Details
Input tokens
$0.20
per 1M tokens
Output tokens
$1.20
per 1M tokens
Estimated cost per 1K requests
$0.80
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.