Skip to content
All systems operational0 AI providers monitored, polled every 2 minutes
Live status

Gemini 3.8 Flash

Mid-tier

by Google

Gemini 3.8 Flash is Google's workhorse Flash release, shipped September 2, 2026 as gemini-3.8-flash. It holds the $0.75 per million input and $3.75 per million output that Gemini 3.7 Flash introduced, with one difference that changes how you should model it: Google printed the expiry. The introductory rate runs through December 31, 2026, and on January 1, 2027 both sides double to $1.50 and $7.50. Batch and Flex are half those rates and Priority is 1.8x. If you are sizing a Flash-tier workload for next year, the number to budget is the post-January one, and if you are choosing between Flash tiers on price alone, the choice you make in September expires in fifteen weeks. The model takes text, image, audio, video, and PDF input with a 1,048,576 token context window and 65,536 output tokens. Artificial Analysis measures GPQA Diamond at 95.3 for the high reasoning tier, which puts it within a point of frontier flagships at roughly a thirteenth of the input price, and Google reports SWE-bench Pro at 61.6 with sharp coding and terminal gains over 3.7 Flash. A separate Cyber variant shipped alongside it behind Fairwind gating and is not part of the standard endpoint.

Input Price

$0.75

per 1M tokens

Output Price

$3.75

per 1M tokens

Context Window

1.0M

tokens

Released

2026-09

API access

Capabilities

textvisionaudiovideotool-usecodereasoning

Key Strengths

  • $0.75/$3.75 through December 31, 2026, doubling to $1.50/$7.50 on January 1, 2027
  • 1,048,576 token context window with 65,536 output tokens
  • Independently measured 95.3 GPQA Diamond at the high reasoning tier
  • Text, image, audio, video, and PDF input on one endpoint
  • Batch and Flex at half rate, Priority at 1.8x

Best For

  • High-volume multimodal processing where cost per token dominates
  • Document and PDF extraction pipelines
  • Coding and terminal agent work on a budget tier
  • Workloads that will be re-priced or re-benchmarked before January 2027

Benchmark Scores

BenchmarkScoreDescription
GPQA Diamond95.3Graduate-level science questions verified by domain experts

Scores sourced from public benchmark datasets. See full benchmark leaderboard for all models.

Pricing Details

Input tokens

$0.75

per 1M tokens

Output tokens

$3.75

per 1M tokens

Estimated cost per 1K requests

$2.62

~1K input + ~500 output tokens avg

Prices are subject to change. Check the official documentation for current pricing. See the cost calculator for detailed estimates.

Related Models

View DocumentationCompare ModelsCost CalculatorFull Pricing Guide