Google DeepMind's most intelligent Flash model, generally available September 2, 2026 and the third Flash-tier release in 43 days. Gemini 3.8 Flash improves on every published benchmark versus Gemini 3.7 Flash, posting 73.7% on DeepSWE v1.1 (Claude Opus 5: 74.0%, GPT-5.6 Sol: 72.7%) and 90.8% on Terminal-Bench 2.1 (3.7 Flash: 81.6%), and is engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. It offers a 1,048,576-token context window, 65,536-token output, and multimodal input (text, image, video, audio, PDF) with text output. API pricing is an introductory $0.75/$3.75 per million tokens, doubling on January 1, 2027.
Parameters
Undisclosed
Context Window
1M
License
Proprietary
Release Date
2026-09-02
API Pricing
Input Price (per 1M tokens)
$0.75
Output Price (per 1M tokens)
$3.75
Billing Mode: standard
Strengths
- •Long-horizon coding near frontier level: DeepSWE v1.1 73.7% vs Claude Opus 5's 74.0%
- •Terminal-Bench 2.1 jumps to 90.8% from 3.7 Flash's 81.6%
- •1M context, 65,536-token output, multimodal input, and three configurable thinking levels
- •Intro price of $0.75/$3.75 through 2026 with default rollout across Gemini surfaces
Weaknesses
- •Intro pricing doubles January 1, 2027 and higher per-task token use can raise real costs ~40%
- •Still trails top models on the hardest agentic tests (Terminal-Bench 4.0-class)
- •No audio/image generation or Live API; computer use remains preview
- •Quiet model-page launch and a ~30-minute listing removal suggest operational roughness
Use Cases
- •Long-horizon software engineering and autonomous coding agents
- •Document-heavy enterprise workflows over a 1M-token context
- •High-volume agentic API workloads with multimodal input
- •Cost-neutral migration off Gemini 3.7 Flash before 2027 pricing doubles
Deep Analysis
Release
September 2, 2026 (GA)
Model ID gemini-3.8-flash; third Flash launch in 43 days
Context Window
1,048,576 tokens
65,536-token output ceiling; text/image/video/audio/PDF input, text output
Intro Price
$0.75 / $3.75 per 1M
Through Dec 31, 2026; doubles to $1.50/$7.50 on Jan 1, 2027; cache inputs ~90% off
DeepSWE v1.1
73.7%
vs Claude Opus 5's 74.0% and GPT-5.6 Sol's 72.7%
Terminal-Bench 2.1
90.8%
Up from 81.6% on Gemini 3.7 Flash
Strengths
- ・Near-frontier long-horizon coding at a budget price: DeepSWE v1.1 73.7% essentially matches Claude Opus 5 (74.0%).
- ・Third Flash release in 43 days means the iteration loop itself is the product — 3.8 Flash inherits a mature, migration-tested line.
- ・1M-token context with true multimodal input (text, image, video, audio, PDF) and three configurable thinking levels.
- ・Intro price of $0.75/$3.75 holds through 2026 and rolls out as the default in the Gemini app, AI Mode, and Sheets for Pro/Ultra.
Weaknesses
- ・Intro pricing doubles on January 1, 2027, and the model consumes more reasoning tokens per task — real job costs can run ~40% higher than 3.7 Flash despite equal per-token prices.
- ・Still trails the top models on the hardest agentic tests (Terminal-Bench 4.0, OSWorld 2.0, GDPVal-AA v2).
- ・No audio or image generation, no Live API; computer use remains in preview.
- ・Google shipped it quietly via a model page rather than an announcement, and briefly pulled the listing for ~30 minutes before re-confirming — operational polish still lags the model quality.
Competitor Comparison
| Model | Arena | SWE | GPQA | Price |
|---|---|---|---|---|
| Gemini 3.8 Flash | 73.7% (DeepSWE v1.1) | 90.8% (Terminal-Bench 2.1) | N/A | $0.75/$3.75 |
| Claude Opus 5 | 74.0% (DeepSWE v1.1) | N/A | N/A | $15/$75 |
| GPT-5.6 Sol | 72.7% (DeepSWE v1.1) | N/A | N/A | $5 input |
| Gemini 3.7 Flash | 65.3% (DeepSWE v1.1) | 81.6% (Terminal-Bench 2.1) | N/A | $0.75/$3.75 |
Gemini 3.8 Flash, generally available September 2, 2026, is Google DeepMind's third Flash-tier release in 43 days and its most intelligent Flash model to date, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. At an intro price of $0.75/$3.75 per million tokens it posts 73.7% on DeepSWE v1.1 — within a point of Claude Opus 5 — making frontier-adjacent agentic coding available at a fraction of the cost.
Sources
Analysis generated: 2026-09-04