Google DeepMindProprietary

Gemini 3.8 Flash

Compare this model

Google DeepMind's most intelligent Flash model, generally available September 2, 2026 and the third Flash-tier release in 43 days. Gemini 3.8 Flash improves on every published benchmark versus Gemini 3.7 Flash, posting 73.7% on DeepSWE v1.1 (Claude Opus 5: 74.0%, GPT-5.6 Sol: 72.7%) and 90.8% on Terminal-Bench 2.1 (3.7 Flash: 81.6%), and is engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. It offers a 1,048,576-token context window, 65,536-token output, and multimodal input (text, image, video, audio, PDF) with text output. API pricing is an introductory $0.75/$3.75 per million tokens, doubling on January 1, 2027.

Parameters

Undisclosed

Context Window

1M

License

Proprietary

Release Date

2026-09-02

API Pricing

Input Price (per 1M tokens)

$0.75

Output Price (per 1M tokens)

$3.75

Billing Mode: standard

Strengths

  • Long-horizon coding near frontier level: DeepSWE v1.1 73.7% vs Claude Opus 5's 74.0%
  • Terminal-Bench 2.1 jumps to 90.8% from 3.7 Flash's 81.6%
  • 1M context, 65,536-token output, multimodal input, and three configurable thinking levels
  • Intro price of $0.75/$3.75 through 2026 with default rollout across Gemini surfaces

Weaknesses

  • Intro pricing doubles January 1, 2027 and higher per-task token use can raise real costs ~40%
  • Still trails top models on the hardest agentic tests (Terminal-Bench 4.0-class)
  • No audio/image generation or Live API; computer use remains preview
  • Quiet model-page launch and a ~30-minute listing removal suggest operational roughness

Use Cases

  • Long-horizon software engineering and autonomous coding agents
  • Document-heavy enterprise workflows over a 1M-token context
  • High-volume agentic API workloads with multimodal input
  • Cost-neutral migration off Gemini 3.7 Flash before 2027 pricing doubles

Deep Analysis

Release

September 2, 2026 (GA)

Model ID gemini-3.8-flash; third Flash launch in 43 days

Context Window

1,048,576 tokens

65,536-token output ceiling; text/image/video/audio/PDF input, text output

Intro Price

$0.75 / $3.75 per 1M

Through Dec 31, 2026; doubles to $1.50/$7.50 on Jan 1, 2027; cache inputs ~90% off

DeepSWE v1.1

73.7%

vs Claude Opus 5's 74.0% and GPT-5.6 Sol's 72.7%

Terminal-Bench 2.1

90.8%

Up from 81.6% on Gemini 3.7 Flash

Strengths

  • Near-frontier long-horizon coding at a budget price: DeepSWE v1.1 73.7% essentially matches Claude Opus 5 (74.0%).
  • Third Flash release in 43 days means the iteration loop itself is the product — 3.8 Flash inherits a mature, migration-tested line.
  • 1M-token context with true multimodal input (text, image, video, audio, PDF) and three configurable thinking levels.
  • Intro price of $0.75/$3.75 holds through 2026 and rolls out as the default in the Gemini app, AI Mode, and Sheets for Pro/Ultra.

Weaknesses

  • Intro pricing doubles on January 1, 2027, and the model consumes more reasoning tokens per task — real job costs can run ~40% higher than 3.7 Flash despite equal per-token prices.
  • Still trails the top models on the hardest agentic tests (Terminal-Bench 4.0, OSWorld 2.0, GDPVal-AA v2).
  • No audio or image generation, no Live API; computer use remains in preview.
  • Google shipped it quietly via a model page rather than an announcement, and briefly pulled the listing for ~30 minutes before re-confirming — operational polish still lags the model quality.

Competitor Comparison

ModelArenaSWEGPQAPrice
Gemini 3.8 Flash73.7% (DeepSWE v1.1)90.8% (Terminal-Bench 2.1)N/A$0.75/$3.75
Claude Opus 574.0% (DeepSWE v1.1)N/AN/A$15/$75
GPT-5.6 Sol72.7% (DeepSWE v1.1)N/AN/A$5 input
Gemini 3.7 Flash65.3% (DeepSWE v1.1)81.6% (Terminal-Bench 2.1)N/A$0.75/$3.75

Gemini 3.8 Flash, generally available September 2, 2026, is Google DeepMind's third Flash-tier release in 43 days and its most intelligent Flash model to date, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. At an intro price of $0.75/$3.75 per million tokens it posts 73.7% on DeepSWE v1.1 — within a point of Claude Opus 5 — making frontier-adjacent agentic coding available at a fraction of the cost.

Analysis generated: 2026-09-04