Qwen3.8-Max is Alibaba's largest and most capable Qwen model to date, a 2.4-trillion-parameter Mixture-of-Experts network that activates just 95B parameters per token. Released August 3, 2026 with open weights (Qwen3.8-2.4T-A95B) following on August 12, it pairs a 1M-token context window with native multimodal input (text, image, video) and $2/$6 per-million-token pricing on QwenCloud. It targets long-horizon agentic coding, autonomous research, and end-to-end product build-out, scoring 74.2% on SWE-bench Verified and 93.0 on PaperBench in Alibaba's evaluations.
Parameters
2.4T (95B active)
Context Window
1M
License
Proprietary
Release Date
2026-08-12
Benchmark Performance
AA Intelligence Index
—
LMArena Elo
—
HLE
—
ARC-AGI-2
—
SWE-bench Verified
74.2
GPQA Diamond
92.6
MMLU-Pro
88.6
LiveCodeBench
—
AIME 2025
—
MATH-500
93.5
API Pricing
Input Price (per 1M tokens)
$2
Output Price (per 1M tokens)
$6
Billing Mode: standard
Strengths
- •First Max-class Qwen model to ship open weights (Qwen3.8-2.4T-A95B), enabling self-hosting at frontier scale
- •2.4T-parameter MoE with only 95B active per token keeps serving cost low relative to dense frontier models
- •1M-token native context with multimodal (text, image, video) input
- •Strong agentic coding: 74.2% SWE-bench Verified, 93.0 PaperBench, 86.6 Terminal-Bench 2.1 (vendor-reported)
Weaknesses
- •Hosted Max API remains proprietary and rate-limited despite the open-weight release
- •Cross-lab benchmark comparisons are directional; some Qwen in-house harnesses are new
- •Datacenter-scale download (~1.6TB class) makes full local deployment impractical for most teams
Use Cases
- •Long-horizon software engineering: autonomous multi-day repository work and PR generation
- •Enterprise knowledge work: financial research, legal compliance audits, and document analysis
- •Multimodal product and design prototyping from text or visual prompts
Deep Analysis
Parameters
2.4T (95B active)
MoE, 512 experts, 10 routed + 1 shared per token
Context Window
1M tokens
262,144 native, extensible to 1,010,000
Input / Output Price
$2 / $6 per 1M
QwenCloud; cached input $0.25/M
SWE-bench Verified
74.2%
Vendor-reported; vs Fable 5 76.8%
Release
Aug 3, 2026 (open weights Aug 12)
First Max-class Qwen open weights
Strengths
- ・First Max-class Qwen model to ship open weights (Qwen3.8-2.4T-A95B), enabling self-hosting at frontier scale
- ・2.4T-parameter MoE activates only 95B per token, keeping serving cost low relative to dense frontier models
- ・1M-token native context with multimodal (text, image, video) input
- ・Strong agentic coding: 74.2% SWE-bench Verified, 93.0 PaperBench, 86.6 Terminal-Bench 2.1 (vendor-reported)
Weaknesses
- ・Hosted Max API remains proprietary and rate-limited despite the open-weight release
- ・Cross-lab benchmark comparisons are directional; some Qwen in-house harnesses are new
- ・Datacenter-scale download (~1.6TB class) makes full local deployment impractical for most teams
Competitor Comparison
| Model | Price |
|---|---|
| Claude Fable 5 (max) | $10/$50 |
| GPT-5.6 Sol (max) | $5/$30 |
| Kimi K3 | $3/$15 |
Qwen3.8-Max is Alibaba's largest and most capable Qwen model, a 2.4-trillion-parameter MoE that activates 95B parameters per token. Released August 3, 2026 with open weights following on August 12, it offers a 1M-token context, native multimodal input, and $2/$6 per-million-token pricing, targeting long-horizon agentic coding and autonomous research.
Sources
Analysis generated: 2026-08-13