Early preview
A

Anthropic

Claude Opus 5.5

Announced Sep 22, 2026

Higher capability at lower operating cost

01

What could it do?

Anthropic introduced an Opus update for coding and professional work.

02

What changed?

Reported Fable 5.1 level performance on many tasks with lower typical costs than Opus 5.

WHY IT MATTERED

Made demanding agent workflows more economical.

03

Where it fell short

Vendor comparisons depend on workload and effort; they are not a universal ranking.

Sources checked Oct 7, 2026

Release facts were checked against the sources below. Performance claims belong to the developers; we have not independently tested these models.

Developer announcement or release log

Benchmark results

EPOCH AI CAPABILITY ESTIMATE
167.3index points
Tested variant: Claude Opus 5.5Source interval: 164.1 to 171.7Variant date in source: Sep 22, 2026

A benchmark estimate, not a percentage or capability multiplier. Reasoning settings are not specified in this source table. Historical estimates can change in later snapshots.

Epoch AI methodology ↗Download the source snapshotChecked Oct 7, 2026 · CC BY 4.0
PUBLISHED BENCHMARK RESULT
58.0index points
Tested: Claude Opus 5.5 (max with fallback)Intelligence Index v4.3.2

Selected highest effort variant where available in the current table. Starred partial results are excluded. Claude fallback configurations can use other models when safeguards intervene.

Source: Artificial Analysis Intelligence Index ↗Download selected resultsChecked Oct 8, 2026
PUBLISHED BENCHMARK RESULT
1507rating points
Tested: claude-opus-5.5-highText Arena Overall · October 8, 2026Reported interval: 1499 to 15156,272 votes

One explicitly named variant per release. Scores come from the same Overall snapshot; preliminary entries and reported intervals are preserved. These are current ratings of earlier variants, not their launch day ratings.

Source: Text Arena ↗Download selected resultsChecked Oct 8, 2026
FOLLOW WHAT HAPPENS NEXT

Breakthroughs, with the followup.

A weekly brief on new discoveries, meaningful checks and what you can actually use.