Google DeepMindプロプライエタリ

Gemini 3.8 Flash

このモデルを比較

Google DeepMindが2026年9月2日に一般提供(GA)したFlash系の最上位モデル。43日間で3回目のFlashリリースで、前モデルGemini 3.7 Flashを全ベンチマークで上回る。DeepSWE v1.1で73.7%(Claude Opus 5は74.0%、GPT-5.6 Solは72.7%)、Terminal-Bench 2.1で90.8%(3.7 Flashは81.6%)を記録し、長期ソフトウェア工学・自律エージェント向けに設計された。100万トークンのコンテキストと65,536トークンの出力上限を持ち、テキスト・画像・動画・音声・PDFのマルチモーダル入力に対応。API価格は100万トークンあたり$0.75/$3.75の導入価格(2027年1月1日から2倍)。

シェア:XはてブLINE

パラメータ

非公開

コンテキスト長

1M

ライセンス

プロプライエタリ

リリース日

2026-09-02

API料金

入力料金(1Mトークンあたり)

$0.75

出力料金(1Mトークンあたり)

$3.75

課金モード: standard

強み

  • DeepSWE v1.1 73.7%とClaude Opus 5(74.0%)に肉薄する長期コーディング性能
  • Terminal-Bench 2.1 90.8%(3.7 Flash 81.6%から大幅向上)
  • 100万トークン・65,536出力・マルチモーダル入力と3段階の思考強度
  • 導入価格$0.75/$3.75で2026年末まで維持、Geminiアプリ等に標準搭載

弱み

  • 2027年1月1日から価格2倍、タスクあたりトークン増で実質コスト約40%上昇の可能性
  • 最高難度のエージェントテスト(Terminal-Bench 4.0等)では最上位モデルに劣る
  • 音声・画像生成とLive API非対応、コンピューター操作はプレビュー段階
  • 発表ではなくモデルページでの静かなリリースと一時的な掲載取り下げで運用面の粗さ

活用例

  • 長期的なソフトウェアエンジニアリングと自律コーディングエージェント
  • 100万トークン文脈のドキュメント処理・エンタープライズ業務
  • マルチモーダル入力が必要な高頻度エージェントAPI
  • Gemini 3.7 Flashからのコスト中立マイグレーション(2026年中)

深度分析

Release

September 2, 2026 (GA)

Model ID gemini-3.8-flash; third Flash launch in 43 days

Context Window

1,048,576 tokens

65,536-token output ceiling; text/image/video/audio/PDF input, text output

Intro Price

$0.75 / $3.75 per 1M

Through Dec 31, 2026; doubles to $1.50/$7.50 on Jan 1, 2027; cache inputs ~90% off

DeepSWE v1.1

73.7%

vs Claude Opus 5's 74.0% and GPT-5.6 Sol's 72.7%

Terminal-Bench 2.1

90.8%

Up from 81.6% on Gemini 3.7 Flash

強み

  • Near-frontier long-horizon coding at a budget price: DeepSWE v1.1 73.7% essentially matches Claude Opus 5 (74.0%).
  • Third Flash release in 43 days means the iteration loop itself is the product — 3.8 Flash inherits a mature, migration-tested line.
  • 1M-token context with true multimodal input (text, image, video, audio, PDF) and three configurable thinking levels.
  • Intro price of $0.75/$3.75 holds through 2026 and rolls out as the default in the Gemini app, AI Mode, and Sheets for Pro/Ultra.

弱み

  • Intro pricing doubles on January 1, 2027, and the model consumes more reasoning tokens per task — real job costs can run ~40% higher than 3.7 Flash despite equal per-token prices.
  • Still trails the top models on the hardest agentic tests (Terminal-Bench 4.0, OSWorld 2.0, GDPVal-AA v2).
  • No audio or image generation, no Live API; computer use remains in preview.
  • Google shipped it quietly via a model page rather than an announcement, and briefly pulled the listing for ~30 minutes before re-confirming — operational polish still lags the model quality.

競合比較

ModelArenaSWEGPQAPrice
Gemini 3.8 Flash73.7% (DeepSWE v1.1)90.8% (Terminal-Bench 2.1)N/A$0.75/$3.75
Claude Opus 574.0% (DeepSWE v1.1)N/AN/A$15/$75
GPT-5.6 Sol72.7% (DeepSWE v1.1)N/AN/A$5 input
Gemini 3.7 Flash65.3% (DeepSWE v1.1)81.6% (Terminal-Bench 2.1)N/A$0.75/$3.75

Gemini 3.8 Flash, generally available September 2, 2026, is Google DeepMind's third Flash-tier release in 43 days and its most intelligent Flash model to date, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows. At an intro price of $0.75/$3.75 per million tokens it posts 73.7% on DeepSWE v1.1 — within a point of Claude Opus 5 — making frontier-adjacent agentic coding available at a fraction of the cost.

分析生成日: 2026-09-04