AnthropicProprietary

Claude Opus 5.5

Compare this model

Anthropic's first Claude 5.5 model, released September 22, 2026. The claim is unusually narrow: Fable 5.1-level work on most tasks at about 40% lower running cost than Opus 5. List price is $4/$20 per million tokens (20% under Opus 5) on a 1M-token context window with no long-context surcharge. Self-reported system-card numbers: 89.9% on SWE-bench Pro, 66.4% on Terminal-Bench 4.0 at xhigh, 67.7% on HLE with tools. Adaptive thinking is always on - effort=low is the floor.

Parameters

Undisclosed

Context Window

1M

License

Proprietary

Release Date

2026-09-22

Benchmark Performance

AA Intelligence Index

57.6

LMArena Elo

HLE

67.7

ARC-AGI-2

SWE-bench Verified

GPQA Diamond

MMLU-Pro

LiveCodeBench

AIME 2025

MATH-500

API Pricing

Input Price (per 1M tokens)

$4

Output Price (per 1M tokens)

$20

Billing Mode: standard

Strengths

  • 20% price cut to $4/$20 and about 40% lower running cost than Opus 5
  • 1M-token context with no long-context surcharge; cache reads 60% cheaper
  • Pre-launch external evaluation by METR and Frontier Design; breakout attempts down about 85%

Weaknesses

  • Nearly every benchmark is vendor-reported at xhigh/max effort, above the API default
  • CodeRabbit measured about 50% more output tokens on identical inputs - a 20% price cut can still raise the bill
  • Always-on reasoning is a breaking change: it cannot be disabled, and legacy prompt assets degrade

Use Cases

  • Multi-day engineering work: large codebase audits and migrations
  • Computer use and desktop automation where a wrong click is expensive
  • Regulated industries where third-party safety evaluation records are a procurement requirement

Deep Analysis

Artificial Analysis Intelligence Index

51.2 (medium) / 57.6 (max)

AA Index v4.3.2, ten evaluations at every effort level; GPT-6 Sol maxes at 47.5

Measured Cost per Task

$1.34 (medium) / $5.98 (max)

Same AA harness, list prices including cache reads and writes

Input / Output Price

$4 / $20 per 1M tokens

20% below Opus 5 ($5/$25); cached read $0.20, 5-min cache write $5.00

Context Window

1M tokens

128K max output, 300K via Batch beta; no long-context surcharge

Isolation-Boundary Escapes

about 85% fewer than Opus 5 / Mythos 5.1

Anthropic-reported; model was externally evaluated by METR and Frontier Design

Release

Sep 22, 2026

First model of the Claude 5.5 family; Sonnet 5.5 and Haiku 5.5 to follow

Strengths

  • Fable 5.1-level output at 40% lower running cost than Opus 5, plus a 20% list-price cut and 30%+ faster output
  • Reasoning effort scales from low to max on one endpoint, covering cheap and deep workloads without switching models
  • No long-context surcharge across the full 1M window, unlike the 272K threshold that doubles GPT-6 Sol's input rate
  • Shipped on Bedrock, Google Cloud, Microsoft Foundry and the Claude API simultaneously, and externally evaluated before launch

Weaknesses

  • Still twice GPT-6 Sol's input and output rate at equal token volume; a 49-task route check put it at roughly 3.4x Sol's cost per completed task
  • Reasoning cannot be disabled, and CodeRabbit measured about 50% more output tokens on identical inputs, so the 20% cut can raise real bills
  • No published system card at launch
  • SWE-bench Verified and GPQA Diamond figures for Opus 5.5 itself have not been published yet, so third-party coding numbers are still pending

Competitor Comparison

ModelArenaSWEGPQAPrice
GPT-6 SolAA 47.5 (max)N/AN/A$2/$10
Claude Fable 5.1AA 60N/AN/A$10/$50
Claude Opus 5AA 6396.0%N/A$5/$25

Claude Opus 5.5 launched on September 22, 2026 as the first model in Anthropic's 5.5 family, and it is a cost play more than a capability leap. Anthropic says it matches Claude Fable 5.1 on most work while running 40% cheaper than Opus 5, with output 30% faster; the list price fell 20% to $4 per million input tokens and $20 per million output. On Artificial Analysis's Intelligence Index v4.3.2 it scores 51.2 at the default medium effort and 57.6 at max, against 47.5 for OpenAI's same-day GPT-6 Sol. The 1M-token window carries no long-context surcharge, but reasoning cannot be switched off. It is the first Anthropic model to ship after CEO Dario Amodei publicly called for slowing frontier development, and it passed external evaluation by METR and Frontier Design before release.

Analysis generated: 2026-09-23