Google DeepMindが2026年9月2日に一般提供(GA)したFlash系の最上位モデル。43日間で3回目のFlashリリースで、前モデルGemini 3.7 Flashを全ベンチマークで上回る。DeepSWE v1.1で73.7%(Claude Opus 5は74.0%、GPT-5.6 Solは72.7%)、Terminal-Bench 2.1で90.8%(3.7 Flashは81.6%)を記録し、長期ソフトウェア工学・自律エージェント向けに設計された。100万トークンのコンテキストと65,536トークンの出力上限を持ち、テキスト・画像・動画・音声・PDFのマルチモーダル入力に対応。API価格は100万トークンあたり$0.75/$3.75の導入価格(2027年1月1日から2倍)。
パラメータ
非公開
コンテキスト長
1M
ライセンス
プロプライエタリ
リリース日
2026-09-02
API料金
入力料金(1Mトークンあたり)
$0.75
出力料金(1Mトークンあたり)
$3.75
課金モード: standard
強み
- •DeepSWE v1.1 73.7%とClaude Opus 5(74.0%)に肉薄する長期コーディング性能
- •Terminal-Bench 2.1 90.8%(3.7 Flash 81.6%から大幅向上)
- •100万トークン・65,536出力・マルチモーダル入力と3段階の思考強度
- •導入価格$0.75/$3.75で2026年末まで維持、Geminiアプリ等に標準搭載
弱み
- •2027年1月1日から価格2倍、タスクあたりトークン増で実質コスト約40%上昇の可能性
- •最高難度のエージェントテスト(Terminal-Bench 4.0等)では最上位モデルに劣る
- •音声・画像生成とLive API非対応、コンピューター操作はプレビュー段階
- •発表ではなくモデルページでの静かなリリースと一時的な掲載取り下げで運用面の粗さ
活用例
- •長期的なソフトウェアエンジニアリングと自律コーディングエージェント
- •100万トークン文脈のドキュメント処理・エンタープライズ業務
- •マルチモーダル入力が必要な高頻度エージェントAPI
- •Gemini 3.7 Flashからのコスト中立マイグレーション(2026年中)
深度分析
Release
September 2, 2026 (GA)
Model ID gemini-3.8-flash; third Flash launch in 43 days
Context Window
1,048,576 tokens
65,536-token output ceiling; text/image/video/audio/PDF input, text output
Intro Price
$0.75 / $3.75 per 1M
Through Dec 31, 2026; doubles to $1.50/$7.50 on Jan 1, 2027; cache inputs ~90% off
DeepSWE v1.1
73.7%
vs Claude Opus 5's 74.0% and GPT-5.6 Sol's 72.7%
Terminal-Bench 2.1
90.8%
Up from 81.6% on Gemini 3.7 Flash
強み
- ・Near-frontier long-horizon coding at a budget price: DeepSWE v1.1 73.7% essentially matches Claude Opus 5 (74.0%).
- ・Third Flash release in 43 days means the iteration loop itself is the product — 3.8 Flash inherits a mature, migration-tested line.
- ・1M-token context with true multimodal input (text, image, video, audio, PDF) and three configurable thinking levels.
- ・Intro price of $0.75/$3.75 holds through 2026 and rolls out as the default in the Gemini app, AI Mode, and Sheets for Pro/Ultra.
弱み
- ・Intro pricing doubles on January 1, 2027, and the model consumes more reasoning tokens per task — real job costs can run ~40% higher than 3.7 Flash despite equal per-token prices.
- ・Still trails the top models on the hardest agentic tests (Terminal-Bench 4.0, OSWorld 2.0, GDPVal-AA v2).
- ・No audio or image generation, no Live API; computer use remains in preview.
- ・Google shipped it quietly via a model page rather than an announcement, and briefly pulled the listing for ~30 minutes before re-confirming — operational polish still lags the model quality.
競合比較
| Model | Arena | SWE | GPQA | Price |
|---|---|---|---|---|
| Gemini 3.8 Flash | 73.7% (DeepSWE v1.1) | 90.8% (Terminal-Bench 2.1) | N/A | $0.75/$3.75 |
| Claude Opus 5 | 74.0% (DeepSWE v1.1) | N/A | N/A | $15/$75 |
| GPT-5.6 Sol | 72.7% (DeepSWE v1.1) | N/A | N/A | $5 input |
| Gemini 3.7 Flash | 65.3% (DeepSWE v1.1) | 81.6% (Terminal-Bench 2.1) | N/A | $0.75/$3.75 |
Gemini 3.8 Flash, generally available September 2, 2026, is Google DeepMind's third Flash-tier release in 43 days and its most intelligent Flash model to date, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. At an intro price of $0.75/$3.75 per million tokens it posts 73.7% on DeepSWE v1.1 — within a point of Claude Opus 5 — making frontier-adjacent agentic coding available at a fraction of the cost.
出典
分析生成日: 2026-09-04