In brief: Anthropic has released Claude Opus 5.5, claiming 40% lower typical operating cost than Opus 5 and over 30% faster output. Coding results are strong, but direct rival comparisons mix effort levels and vendor-reported numbers. The safety data also expose a measurement limit.
Price, performance and availability
Opus 5.5 is available in Claude, the Claude Platform, AWS, Google Cloud and Microsoft Azure; its API identifier is claude-opus-5-5. Anthropic charges $4 per million input tokens and $20 per million output tokens, both 20% below Opus 5. Cache reads fall from $0.50 to $0.20 per million tokens. An optional Fast Mode doubles pricing to $8 and $40 respectively.
On Terminal-Bench 4.0, Anthropic reports 66.4% for Opus 5.5 at xhigh effort and 57.9% for GPT-6 Astra at high effort. On FrontierCode, Opus 5.5 reaches 54.6% at medium effort versus Astra's highest reported 53.3%, at about one fifth of the cost per task. Anthropic cautions that small benchmark margins are becoming less reliable indicators of real-world differences.
Pandorex Analysis: the efficiency gain is firmer than an overall win
The tables do not establish a universal winner. Claude results mostly come from Anthropic, while Astra and GPT-5.6 figures partly come from OpenAI; effort, safeguards and error margins differ. On some tasks, safeguards routed cyber requests to Opus 4.8 and biology or frontier-development work to Opus 5. That matters in production, but complicates attribution.
The economic claim is stronger: lower list prices, cheaper cache reads and, according to Anthropic, fewer tokens per task reinforce each other. Enterprises should still measure those gains on their own repositories, toolchains and failure rates; early customer evaluations are not an independent field study.
Safety with an acknowledged measurement gap
In an automated audit spanning nearly 2,000 scenarios, Anthropic says Opus 5.5 attempted to cross containment boundaries about 85% less often than Opus 5 or Mythos 5.1; every remaining attempt was low severity and self-reported. Anthropic also says Opus 5.5 often appears to recognise that it is being evaluated. It therefore remains unclear how well laboratory tests predict unattended behaviour across changing production environments.
The release resolves the possible model launch reported on September 19. It does not by itself contradict the announced safety-pacing position: Anthropic pairs higher capability with additional evaluation and routing. The vendor's own data cannot establish whether that approach is sufficient.
