Gemini 3.7 Flash
Mid-tierby Google
Gemini 3.7 Flash went generally available on August 13, 2026, three weeks after Gemini 3.6 Flash and with Gemini 3.5 Pro still unshipped. The model id is gemini-3.7-flash. Introductory pricing is $0.75 per million input tokens and $3.75 per million output through December 31, 2026, exactly half the 3.6 Flash rate, after which it reverts to $1.50/$7.50. Batch and Flex processing halve that again to $0.375 and $1.875. It carries a 1,048,576 token input window, up to 65,536 output tokens, and accepts text, image, audio, video, and PDF input. Google says it did not retrain from scratch: 3.7 Flash is built on algorithmic improvements and user feedback applied to the 3.6 line, and it fully replaces rather than supplements the previous version. The published gains are concentrated in coding and agent work. DeepSWE v1.1 moves from 48.6 to 65.3 percent, FrontierCode 1.1 Main from 34.4 to 43.6, Terminal-Bench 2.1 lands at 85.8, and GDM-MRCR v2 reports 97.0 percent long-context retrieval at 128K. The honest caveat is the same one that applied to 3.6: these are Google-reported figures, Artificial Analysis has no SWE-bench Verified score for 3.7 Flash yet, and independent runs have not landed. The strategic read matters as much as the numbers. Google is shipping its cheap tier on a three-week cadence at falling prices while its flagship Pro model slips a third and fourth time, which makes Flash the practical Google model for anyone building agents today. Day-one availability spans the Gemini API, Google AI Studio, Android Studio, Google Antigravity, the Gemini Enterprise Agent Platform, and the Gemini app.
Input Price
$0.75
per 1M tokens
Output Price
$3.75
per 1M tokens
Context Window
1.0M
tokens
Released
2026-08
API access
Capabilities
Key Strengths
- ✓$0.75/$3.75 introductory pricing, half the Gemini 3.6 Flash rate
- ✓DeepSWE v1.1 up from 48.6 to 65.3 percent on Google's own numbers
- ✓85.8 percent on Terminal-bench 2.1 for agent workflows
- ✓1,048,576 token input window with 64K max output
- ✓97.0 percent GDM-MRCR v2 long-context retrieval at 128K
- ✓Text, image, audio, video, and PDF input
Best For
- ▸Agentic coding and multi-step tool use at low cost
- ▸High-volume web development and code review
- ▸Long-context retrieval over large document sets
- ▸Agent pipelines where per-task token spend dominates
Pricing Details
Input tokens
$0.75
per 1M tokens
Output tokens
$3.75
per 1M tokens
Estimated cost per 1K requests
$2.62
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.
Related Models
Gemini 2.5 Pro
Flagship$1.25 in / $10.00 out
Gemini 2.0 Flash
Budget$0.10 in / $0.40 out
Gemini 3.1 Flash-Lite
Budget$0.25 in / $1.50 out
Gemini 3.5 Flash
Mid-tier$1.50 in / $9.00 out
Gemini 3.6 Flash
Mid-tier$1.50 in / $7.50 out
Gemini 3.5 Flash-Lite
Budget$0.30 in / $2.50 out
Claude Opus 4.7
Mid-tierAnthropic
$15.00 in / $75.00 out
Claude Sonnet 5
Mid-tierAnthropic
$2.00 in / $10.00 out
Claude Sonnet 4.6
Mid-tierAnthropic
$3.00 in / $15.00 out
GPT-5.6 Terra
Mid-tierOpenAI
$2.00 in / $12.00 out