ModelsAnnouncement

Claude Opus 5.5 arrives

Claude Opus 5.5 begins Anthropic’s new model family, with an emphasis on complex coding and knowledge work. The release pairs capability claims with external evaluation and additional safeguards.

Complex work still needs a review boundary

The Opus 5.5 announcement describes testing by external evaluators, including METR, and stronger results on Anthropic’s automated behavioral audit. It also reports lower running costs than Opus 5 and improvements on demanding software tasks.

Anthropic released Claude Opus 5.5 as the first model in its new family, targeting complex coding and knowledge work. The announcement describes external evaluation, new safeguards and lower pricing, with specialized verification for some research uses.
Trion

Complex work still needs review

Report boundary
The announcement’s evaluations do not establish success on an unrelated production task.
The announcement’s evaluations do not establish success on an unrelated production task.
View data
EvidenceMeaning
Report boundaryThe announcement’s evaluations do not establish success on an unrelated production task.

Anthropic · published 2026-09-22. Source-bound illustration, not a performance benchmark.

Download image

Anthropic discusses safeguards for biology and cybersecurity, with verified access for particular research uses. The company’s tests and customer examples are useful release evidence, but do not demonstrate that the model will succeed on an unrelated production task or eliminate the need for controls.

Original sources

  1. Anthropic: original announcementwww.anthropic.com

Checked 9 Oct 2026 · A manually curated edition. Availability may change; company performance claims are not Trion test results. Editorial method.