AlibabaProprietary

Qwen3.8-Max

Compare this model

Qwen3.8-Max is Alibaba's largest and most capable Qwen model to date, a 2.4-trillion-parameter Mixture-of-Experts network that activates just 95B parameters per token. Released August 3, 2026 with open weights (Qwen3.8-2.4T-A95B) following on August 12, it pairs a 1M-token context window with native multimodal input (text, image, video) and $2/$6 per-million-token pricing on QwenCloud. It targets long-horizon agentic coding, autonomous research, and end-to-end product build-out, scoring 74.2% on SWE-bench Verified and 93.0 on PaperBench in Alibaba's evaluations.

Parameters

2.4T (95B active)

Context Window

1M

License

Proprietary

Release Date

2026-08-12

Benchmark Performance

AA Intelligence Index

LMArena Elo

HLE

ARC-AGI-2

SWE-bench Verified

74.2

GPQA Diamond

92.6

MMLU-Pro

88.6

LiveCodeBench

AIME 2025

MATH-500

93.5

API Pricing

Input Price (per 1M tokens)

$2

Output Price (per 1M tokens)

$6

Billing Mode: standard

Strengths

  • First Max-class Qwen model to ship open weights (Qwen3.8-2.4T-A95B), enabling self-hosting at frontier scale
  • 2.4T-parameter MoE with only 95B active per token keeps serving cost low relative to dense frontier models
  • 1M-token native context with multimodal (text, image, video) input
  • Strong agentic coding: 74.2% SWE-bench Verified, 93.0 PaperBench, 86.6 Terminal-Bench 2.1 (vendor-reported)

Weaknesses

  • Hosted Max API remains proprietary and rate-limited despite the open-weight release
  • Cross-lab benchmark comparisons are directional; some Qwen in-house harnesses are new
  • Datacenter-scale download (~1.6TB class) makes full local deployment impractical for most teams

Use Cases

  • Long-horizon software engineering: autonomous multi-day repository work and PR generation
  • Enterprise knowledge work: financial research, legal compliance audits, and document analysis
  • Multimodal product and design prototyping from text or visual prompts

Deep Analysis

Parameters

2.4T (95B active)

MoE, 512 experts, 10 routed + 1 shared per token

Context Window

1M tokens

262,144 native, extensible to 1,010,000

Input / Output Price

$2 / $6 per 1M

QwenCloud; cached input $0.25/M

SWE-bench Verified

74.2%

Vendor-reported; vs Fable 5 76.8%

Release

Aug 3, 2026 (open weights Aug 12)

First Max-class Qwen open weights

Strengths

  • First Max-class Qwen model to ship open weights (Qwen3.8-2.4T-A95B), enabling self-hosting at frontier scale
  • 2.4T-parameter MoE activates only 95B per token, keeping serving cost low relative to dense frontier models
  • 1M-token native context with multimodal (text, image, video) input
  • Strong agentic coding: 74.2% SWE-bench Verified, 93.0 PaperBench, 86.6 Terminal-Bench 2.1 (vendor-reported)

Weaknesses

  • Hosted Max API remains proprietary and rate-limited despite the open-weight release
  • Cross-lab benchmark comparisons are directional; some Qwen in-house harnesses are new
  • Datacenter-scale download (~1.6TB class) makes full local deployment impractical for most teams

Competitor Comparison

ModelPrice
Claude Fable 5 (max)$10/$50
GPT-5.6 Sol (max)$5/$30
Kimi K3$3/$15

Qwen3.8-Max is Alibaba's largest and most capable Qwen model, a 2.4-trillion-parameter MoE that activates 95B parameters per token. Released August 3, 2026 with open weights following on August 12, it offers a 1M-token context, native multimodal input, and $2/$6 per-million-token pricing, targeting long-horizon agentic coding and autonomous research.

Analysis generated: 2026-08-13