Gemini 3.8 Flash
Mid-tierby Google
Gemini 3.8 Flash is Google's workhorse Flash release, shipped September 2, 2026 as gemini-3.8-flash. It holds the $0.75 per million input and $3.75 per million output that Gemini 3.7 Flash introduced, with one difference that changes how you should model it: Google printed the expiry. The introductory rate runs through December 31, 2026, and on January 1, 2027 both sides double to $1.50 and $7.50. Batch and Flex are half those rates and Priority is 1.8x. If you are sizing a Flash-tier workload for next year, the number to budget is the post-January one, and if you are choosing between Flash tiers on price alone, the choice you make in September expires in fifteen weeks. The model takes text, image, audio, video, and PDF input with a 1,048,576 token context window and 65,536 output tokens. Artificial Analysis measures GPQA Diamond at 95.3 for the high reasoning tier, which puts it within a point of frontier flagships at roughly a thirteenth of the input price, and Google reports SWE-bench Pro at 61.6 with sharp coding and terminal gains over 3.7 Flash. A separate Cyber variant shipped alongside it behind Fairwind gating and is not part of the standard endpoint.
Input Price
$0.75
per 1M tokens
Output Price
$3.75
per 1M tokens
Context Window
1.0M
tokens
Released
2026-09
API access
Capabilities
Key Strengths
- ✓$0.75/$3.75 through December 31, 2026, doubling to $1.50/$7.50 on January 1, 2027
- ✓1,048,576 token context window with 65,536 output tokens
- ✓Independently measured 95.3 GPQA Diamond at the high reasoning tier
- ✓Text, image, audio, video, and PDF input on one endpoint
- ✓Batch and Flex at half rate, Priority at 1.8x
Best For
- ▸High-volume multimodal processing where cost per token dominates
- ▸Document and PDF extraction pipelines
- ▸Coding and terminal agent work on a budget tier
- ▸Workloads that will be re-priced or re-benchmarked before January 2027
Benchmark Scores
| Benchmark | Score | Description |
|---|---|---|
| GPQA Diamond | 95.3 | Graduate-level science questions verified by domain experts |
Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.
Pricing Details
Input tokens
$0.75
per 1M tokens
Output tokens
$3.75
per 1M tokens
Estimated cost per 1K requests
$2.62
~1K input + ~500 output tokens avg
Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.
Related Models
Gemini 2.5 Pro
Flagship$1.25 in / $10.00 out
Gemini 2.0 Flash
Budget$0.10 in / $0.40 out
Gemini 3.1 Flash-Lite
Budget$0.25 in / $1.50 out
Gemini 3.5 Flash
Mid-tier$1.50 in / $9.00 out
Gemini 3.7 Flash
Mid-tier$0.75 in / $3.75 out
Gemini 3.6 Flash
Mid-tier$1.50 in / $7.50 out
Gemini 3.5 Flash-Lite
Budget$0.30 in / $2.50 out
Claude Opus 4.7
Mid-tierAnthropic
$5.00 in / $25.00 out
Claude Sonnet 5
Mid-tierAnthropic
$2.00 in / $10.00 out
Claude Sonnet 4.6
Mid-tierAnthropic
$3.00 in / $15.00 out
GPT-5.6 Terra
Mid-tierOpenAI
$2.00 in / $12.00 out