Gemini 3.6 Flash
Mid-tierby Google
Gemini 3.6 Flash went generally available on July 21, 2026, and the pitch is efficiency rather than raw intelligence. The model id is gemini-3.6-flash, priced at $1.50 per million input tokens and $7.50 per million output, a cut from the $9.00 output rate on Gemini 3.5 Flash, with cached input at $0.15. It carries a 1,048,576 token input window, up to 65,536 output tokens, and takes text, image, audio, video, and PDF input. Google says it is built directly on Gemini 3.5 Flash and consumes 17 percent fewer output tokens on the Artificial Analysis Index, with up to 65 percent fewer on DeepSWE. Google's published gains over 3.5 Flash are DeepSWE 49 percent against 37, MLE-Bench 63.9 against 49.7, OSWorld-Verified 83.0 against 78.4, and GDPval-AA v2 at 1421 Elo against 1349. The honest caveat is that Artificial Analysis scored the Intelligence Index flat at 50, identical to 3.5 Flash, while measuring average time per task falling from 2.7 minutes to 1.3 and measured cost per task from $0.59 to $0.50. Read it as a per-task economics upgrade for existing Flash workloads, not a capability jump. Computer use is now a built-in client-side tool through the Gemini API, and the knowledge cutoff is March 2026. Day-one availability spans Google AI Studio, the Gemini API, Android Studio, Google Antigravity, the Gemini Enterprise Agent Platform, and the Gemini app.
Input Price
$1.50
per 1M tokens
Output Price
$7.50
per 1M tokens
Context Window
1.0M
tokens
Released
2026-07
API access
Capabilities
Key Strengths
- ✓$7.50 per 1M output, down from $9.00 on 3.5 Flash
- ✓17 percent fewer output tokens on the Artificial Analysis Index
- ✓Measured time per task roughly halved, 2.7 minutes to 1.3
- ✓OSWorld-Verified 83.0 percent with computer use as a built-in tool
- ✓1M token input window with 64K output
- ✓Text, image, audio, video, and PDF input
Best For
- ▸Long-running agentic workflows where token spend dominates
- ▸Agentic coding and multi-step tool use
- ▸Computer-use and GUI automation
- ▸Document parsing, chart analysis, and report drafting
Pricing Details
Input tokens
$1.50
per 1M tokens
Output tokens
$7.50
per 1M tokens
Estimated cost per 1K requests
$5.25
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.
Related Models
Gemini 2.5 Pro
Flagship$1.25 in / $10.00 out
Gemini 2.0 Flash
Budget$0.10 in / $0.40 out
Gemini 3.1 Flash-Lite
Budget$0.25 in / $1.50 out
Gemini 3.5 Flash
Mid-tier$1.50 in / $9.00 out
Gemini 3.5 Flash-Lite
Budget$0.30 in / $2.50 out
Claude Opus 4.7
Mid-tierAnthropic
$15.00 in / $75.00 out
Claude Sonnet 5
Mid-tierAnthropic
$2.00 in / $10.00 out
Claude Sonnet 4.6
Mid-tierAnthropic
$3.00 in / $15.00 out
GPT-5.6 Terra
Mid-tierOpenAI
$2.00 in / $12.00 out