Anthropic's first Claude 5.5 model, released September 22, 2026. The claim is unusually narrow: Fable 5.1-level work on most tasks at about 40% lower running cost than Opus 5. List price is $4/$20 per million tokens (20% under Opus 5) on a 1M-token context window with no long-context surcharge. Self-reported system-card numbers: 89.9% on SWE-bench Pro, 66.4% on Terminal-Bench 4.0 at xhigh, 67.7% on HLE with tools. Adaptive thinking is always on - effort=low is the floor.
Parameters
Undisclosed
Context Window
1M
License
Proprietary
Release Date
2026-09-22
Benchmark Performance
AA Intelligence Index
57.6
LMArena Elo
—
HLE
67.7
ARC-AGI-2
—
SWE-bench Verified
—
GPQA Diamond
—
MMLU-Pro
—
LiveCodeBench
—
AIME 2025
—
MATH-500
—
API Pricing
Input Price (per 1M tokens)
$4
Output Price (per 1M tokens)
$20
Billing Mode: standard
Strengths
- •20% price cut to $4/$20 and about 40% lower running cost than Opus 5
- •1M-token context with no long-context surcharge; cache reads 60% cheaper
- •Pre-launch external evaluation by METR and Frontier Design; breakout attempts down about 85%
Weaknesses
- •Nearly every benchmark is vendor-reported at xhigh/max effort, above the API default
- •CodeRabbit measured about 50% more output tokens on identical inputs - a 20% price cut can still raise the bill
- •Always-on reasoning is a breaking change: it cannot be disabled, and legacy prompt assets degrade
Use Cases
- •Multi-day engineering work: large codebase audits and migrations
- •Computer use and desktop automation where a wrong click is expensive
- •Regulated industries where third-party safety evaluation records are a procurement requirement
Deep Analysis
Artificial Analysis Intelligence Index
51.2 (medium) / 57.6 (max)
AA Index v4.3.2, ten evaluations at every effort level; GPT-6 Sol maxes at 47.5
Measured Cost per Task
$1.34 (medium) / $5.98 (max)
Same AA harness, list prices including cache reads and writes
Input / Output Price
$4 / $20 per 1M tokens
20% below Opus 5 ($5/$25); cached read $0.20, 5-min cache write $5.00
Context Window
1M tokens
128K max output, 300K via Batch beta; no long-context surcharge
Isolation-Boundary Escapes
about 85% fewer than Opus 5 / Mythos 5.1
Anthropic-reported; model was externally evaluated by METR and Frontier Design
Release
Sep 22, 2026
First model of the Claude 5.5 family; Sonnet 5.5 and Haiku 5.5 to follow
Strengths
- ・Fable 5.1-level output at 40% lower running cost than Opus 5, plus a 20% list-price cut and 30%+ faster output
- ・Reasoning effort scales from low to max on one endpoint, covering cheap and deep workloads without switching models
- ・No long-context surcharge across the full 1M window, unlike the 272K threshold that doubles GPT-6 Sol's input rate
- ・Shipped on Bedrock, Google Cloud, Microsoft Foundry and the Claude API simultaneously, and externally evaluated before launch
Weaknesses
- ・Still twice GPT-6 Sol's input and output rate at equal token volume; a 49-task route check put it at roughly 3.4x Sol's cost per completed task
- ・Reasoning cannot be disabled, and CodeRabbit measured about 50% more output tokens on identical inputs, so the 20% cut can raise real bills
- ・No published system card at launch
- ・SWE-bench Verified and GPQA Diamond figures for Opus 5.5 itself have not been published yet, so third-party coding numbers are still pending
Competitor Comparison
| Model | Arena | SWE | GPQA | Price |
|---|---|---|---|---|
| GPT-6 Sol | AA 47.5 (max) | N/A | N/A | $2/$10 |
| Claude Fable 5.1 | AA 60 | N/A | N/A | $10/$50 |
| Claude Opus 5 | AA 63 | 96.0% | N/A | $5/$25 |
Claude Opus 5.5 launched on September 22, 2026 as the first model in Anthropic's 5.5 family, and it is a cost play more than a capability leap. Anthropic says it matches Claude Fable 5.1 on most work while running 40% cheaper than Opus 5, with output 30% faster; the list price fell 20% to $4 per million input tokens and $20 per million output. On Artificial Analysis's Intelligence Index v4.3.2 it scores 51.2 at the default medium effort and 57.6 at max, against 47.5 for OpenAI's same-day GPT-6 Sol. The 1M-token window carries no long-context surcharge, but reasoning cannot be switched off. It is the first Anthropic model to ship after CEO Dario Amodei publicly called for slowing frontier development, and it passed external evaluation by METR and Frontier Design before release.
Sources
Analysis generated: 2026-09-23