Gemini 3.8 Flash
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.
- Capability
- 156.9ECI · #15 of 148
- Input
- $0.75per 1M tokens
- Output
- $3.75per 1M tokens
- Context
- 1.05M66K max output
What it costs
- Input
- $0.75per million tokens
- Output
- $3.75per million tokens
- Cached input
- $0.075per million tokens
- Blended (3:1)
- $1.50pricier than 64% of priced models
Typical monthly bills
| Monthly usage | Estimated cost |
|---|---|
| Side project2M input + 0.5M output tokens | $3.38 |
| Team assistant25M input + 5M output tokens | $37.50 |
| Production app250M input + 50M output tokens | $375.00 |
Official Google API price, as listed on models.dev.
Independent benchmarks
#15 of 148 scored models
The shaded band is Epoch AI’s confidence range (154.6–160.4); the tick marks the median scored model.
Scores from Epoch AI, run independently of Google.
The details
- Lab
- Released
- Sep 2, 2026
- Context window
- 1,048,576 tokens
- Max output
- 65,536 tokens
- Inputs
- Text, Images, PDFs, Audio, Video
- Output
- Text
- Reasoning
- Adjustable effort low · medium · high
- Tool calling
- Yes
- Structured output
- Yes
- Weights
- Proprietary
- API model ID
gemini-3.8-flashon Google- Availability
- 21 API providerslisted on models.dev
What Google claims
Published by the lab at launch. Settings vary, so compare these only with care.
| Benchmark | Score | Setting | Source |
|---|---|---|---|
| DeepSWE v1.1 | 73.7 | high thinking | Source |
| GDPval-AA v2 | 1545 Elo | — | Source |
| Finance Agent v2 | 61.4 | — | Source |
| Harvey Legal Agent Benchmark | 10 | — | Source |
| Terminal-Bench v2.1 | 89.4 | — | Source |
| Terminal-Bench v4.0 | 19.1 | — | Source |
| GDP.PDF | 35 | — | Source |
| CharXiv Reasoning | 86.2 | without tools | Source |
| LVBench | 87.8 | agentic | Source |
| LVBench | 87.1 | static; 1024 frames; without tools | Source |
| HLE-Verified | 54.9 | — | Source |
| OSWorld v2.0 | 59 | batched tool calls; best of 3 runs; 500 steps | Source |
More from Google
About Gemini 3.8 Flash
How much does Gemini 3.8 Flash cost?
Gemini 3.8 Flash costs $0.75 input / $3.75 output per million tokens (official Google API price). Cached input is $0.075 per million tokens. At a 3:1 input-to-output mix that is $1.50 per million tokens, more expensive than 64% of the 360 priced models we track.
What is the context window of Gemini 3.8 Flash?
Gemini 3.8 Flash accepts up to 1,048,576 tokens per request and can write up to 65,536 tokens in one response.
How good is Gemini 3.8 Flash?
Epoch AI gives Gemini 3.8 Flash a Capabilities Index score of 156.9 (likely range 154.6–160.4), ranking it #15 of 148 models Epoch has scored. Epoch AI benchmark results: GPQA Diamond 95.4%, FrontierMath Tiers 1–3 68.4%, OTIS Mock AIME 2024–2025 98.9%, SimpleQA Verified 69.7%.
Is Gemini 3.8 Flash open source?
No. Gemini 3.8 Flash is proprietary; you use it through Google’s API or partner platforms.
What inputs does Gemini 3.8 Flash support?
Gemini 3.8 Flash accepts text, images, PDFs, audio and video and replies in text. It is a reasoning model with low, medium and high effort settings, supports tool calling and can return structured JSON output.
When was Gemini 3.8 Flash released?
Google released Gemini 3.8 Flash on Sep 2, 2026.
What are the best alternatives to Gemini 3.8 Flash?
The closest current models from other labs on capability, price and release date are Muse Spark 1.3 (Meta, ECI 156.9, $1.25 / $4.25), GLM-5.3 (Z.ai (Zhipu), ECI 155.8, $1.40 / $4.40), DeepSeek V4 Pro 0813 (DeepSeek, ECI 155.4, $0.66 / $1.98) and Grok 4.6 (xAI, ECI 156.6, $2.00 / $6.00).