Gemini 3.8 Flash
Gemini 3.8 Flash is a Google model with 1.05M context, $0.75 in / $3.75 out per million tokens, last verified September 3, 2026. Open weights: no. Public-board scores are linked to the publisher; vendor-only figures stay in the watch-out, not on the board.
Gemini 3.8 Flash is Google's most-intelligent workhorse Flash — tied for the top of DeepSWE 1.1 with Claude Opus 5 at roughly a fifth of the cost per task, three Flash releases in six weeks.
Catalog record checked September 3, 2026Individual provider fields may change
| Specification | Gemini 3.8 Flash |
|---|---|
| Provider | |
| Tier | Balanced |
| Context window | 1.05M |
| Max output | 66K |
| Input / 1M tokens | $0.75 |
| Output / 1M tokens | $3.75 |
| Weights | Closed |
| Parameters | Not disclosed |
| Reasoning levels | low, medium, high |
| Modalities | text, image, video, audio, pdf |
| API model id | gemini-3.8-flash |
| Released | September 2, 2026 |
Pricing tiers: Intro $0.75/$3.75 per MTok through 2026-12-31, rising to $1.50/$7.50 from 2027-01-01; batch/Flex half price. Free tier available. Companion Gemini 3.8 Flash Cyber (defensive security) is restricted to the Fairwind Program with no public pricing.
Verified evidence
Published benchmark results
Each result keeps its source and measurement date visible. A missing benchmark is not treated as a zero.
| Benchmark | Score | Measured | Source |
|---|---|---|---|
| Artificial Analysis Intelligence Index [high] | 58.7 | 2026-09-02 | Artificial Analysis · View source |
| Terminal-Bench 2.1 | 89.4 | 2026-09-02 | Google launch eval table (transcribed by Vellum) · View source |
| Humanity's Last Exam | 54.9 | 2026-09-02 | Google official blog (vendor, HLE-Verified full set) · View source |
Benchmark
DeepSWE 1.1 in context
The full local snapshot puts this model beside the wider field, including cost per completed task.
Local leader
Gemini 3.8 Flash [high]
74%
Rows shown
27
Highest published reasoning effort per model (not best Pass@1)
Snapshot date
2026-09-03
Mirrored from deepswe.datacurve.ai
Better is toward the top-right (higher pass rate, lower cost). X-axis is reversed to match DeepSWE’s public chart. v1.1 uses average cost / tokens / steps; v1 uses published medians.
| # | Model | Pass@1 | Cost / task | Tokens / task | Steps / task |
|---|---|---|---|---|---|
| 1 | Gemini 3.8 Flash [high] | 74% | $2.36 | 143k | 166 |
| 2 | Claude Opus 5 [max] | 74% | $11.84 | 118k | 99 |
| 3 | GPT-5.6 Sol [max] | 73% | $8.39 | 60k | 61 |
| 4 | Claude Fable 5 [max] | 70% | $21.63 | 119k | 88 |
| 5 | GPT-5.6 Terra [max] | 70% | $4.95 | 72k | 76 |
| 6 | GLM 5.3 [max] | 69% | $3.99 | 80k | 124 |
| 7 | Kimi K3 [max] | 69% | $4.65 | 82k | 98 |
| 8 | GPT-5.6 Luna [max] | 67% | $3.03 | 73k | 102 |
| 9 | GPT-5.5 [xhigh] | 67% | $7.23 | 46k | 82 |
| 10 | Grok 4.6 [xhigh] | 67% | $5.50 | 71k | 87 |
| 11 | Gemini 3.7 Flash [high] | 65% | $2.18 | 107k | 125 |
| 12 | GLM 5.3 Flash [max] | 63% | $0.48 | 73k | 123 |
| 13 | DeepSeek V4-Pro [max] | 63% | $0.24 | 106k | 155 |
| 14 | Claude Opus 4.8 [max] | 59% | $13.22 | 135k | 120 |
| 15 | Qwen3.8-Max [xhigh] | 58% | $3.73 | 95k | 111 |
| 16 | Muse Spark 1.2 [xhigh] | 55% | $3.70 | 99k | 101 |
| 17 | Claude Sonnet 5 [max] | 54% | $26.40 | 214k | 268 |
| 18 | Grok 4.5 [high] | 54% | $2.42 | 36k | 61 |
| 19 | DeepSeek V4-Flash [max] | 53% | $0.10 | 108k | 153 |
| 20 | Muse Spark 1.1 [xhigh] | 53% | $2.36 | 74k | 96 |
| 21 | GPT-5.4 [xhigh] | 52% | $5.65 | 71k | 70 |
| 22 | Gemini 3.6 Flash [high] | 47% | $4.42 | 96k | 117 |
| 23 | GLM 5.2 [max] | 44% | $3.92 | 78k | 129 |
| 24 | Gemini 3.5 Flash [high] | 36% | $3.45 | 76k | 105 |
| 25 | Kimi K2.7 Code | 31% | $2.82 | 59k | 149 |
| 26 | Claude Sonnet 4.6 [high] | 30% | $5.52 | 76k | 134 |
| 27 | Gemini 3.1 Pro [high] | 12% | $2.14 | 28k | 76 |
DeepSWE “Best” picks the highest published reasoning effort per model (not the highest pass rate). Small gaps may not be statistically meaningful — confirm on deepswe.datacurve.ai.
Best for
- Agentic coding at scale
- Mid-difficulty engineering
- Multimodal pipelines
Watch out
History
Gemini 3.8 Flash timeline
Release and status events for this model that passed two-source verification — dated, sourced, and linked.
- Gemini 3.8 Flash + Flash CyberReleaseSeptember 2, 2026
Google unveils Gemini 3.8 Flash workhorse with 90.8% on Terminal-Bench 2.1 and companion defensive cyber edition via the Fairwind Program.
Google DeepMind — Gemini 3.8 Flash announcement
Source receipts
Catalog figures for this model were checked against the following sources.
- Google — Gemini 3.8 Flash and 3.8 Flash Cyber (accessed 2026-09-03)
- Gemini API pricing (gemini-3.8-flash $0.75/$3.75 intro) (accessed 2026-09-03)
- DeepSWE 1.1 leaderboard — gemini-3.8-flash [high] 73.8% at $2.36/task (accessed 2026-09-03)
Common questions
Gemini 3.8 Flash
Answered from the verified figures on this page rather than general guidance.
How much does Gemini 3.8 Flash cost per million tokens?
Gemini 3.8 Flash is listed at $0.75 per million input tokens and $3.75 per million output tokens at standard rates. Output tokens usually dominate real bills, so weigh the output rate more heavily than the input rate. Intro $0.75/$3.75 per MTok through 2026-12-31, rising to $1.50/$7.50 from 2027-01-01; batch/Flex half price. Free tier available. Companion Gemini 3.8 Flash Cyber (defensive security) is restricted to the Fairwind Program with no public pricing.
What is Gemini 3.8 Flash's context window?
Gemini 3.8 Flash accepts about 1.05M tokens of context. That only matters if you routinely send very long documents, large codebases, or multi-turn histories that approach that limit.
What is Gemini 3.8 Flash best for?
Gemini 3.8 Flash is a balanced tier from Google. It suits agentic coding at scale, mid-difficulty engineering, multimodal pipelines. Intro pricing doubles on 2027-01-01, and Google says it can consume more tokens than 3.7 Flash — 3.7 Flash stays available for efficiency-first workloads.
Compare it
Featured head-to-head comparisons
A deliberately selected pairing for this model, rather than an automatically generated matrix.