Claude Opus 5.5: Benchmarks, Pricing, and What Changed
Claude Opus 5.5 Guide
Claude Opus 5.5 is Anthropic’s recommended default model, released September 22, 2026. It costs 20% less per token than Opus 5 ($4/$20 per million tokens), outscores Claude Fable 5.1 on every benchmark in Anthropic’s launch table, and is now the default model in Claude Code on every paid plan.
Quick answer: Claude Opus 5.5 (claude-opus-5-5) is the model Anthropic now tells you to start with for most workloads. It’s priced at $4 input / $20 output per million tokens, runs more than 30% faster than Opus 5, and Anthropic estimates it costs about 40% less on typical workloads because it also uses fewer tokens to finish. On Anthropic’s benchmarks it beats the pricier Claude Fable 5.1 across the board — 66.4% vs 55.8% on Terminal-Bench 4.0, for example — though Anthropic reserves Fable 5.1 for work where Opus 5.5 at higher effort still falls short.
What’s New in Claude Opus 5.5
Opus 5.5 succeeds Claude Opus 5, which is now a legacy model (still available, retirement no sooner than July 24, 2027). Anthropic built it for long-running agentic coding and knowledge work, and the headline changes are about doing the same work with less:
- Cheaper per token: $4/$20 per million tokens versus Opus 5’s $5/$25 — 20% off input and output. Cache reads drop 60%, from $0.50 to $0.20.
- Fewer tokens per task: Anthropic says typical workloads cost about 40% less than on Opus 5 once token efficiency is counted. Customers report the same pattern: Box measured a third of the tokens Opus 5 used, and Optiver cut the cost of one workload by 40–50%.
- Faster: output generation is more than 30% faster than Opus 5, and a new fast mode (research preview) runs up to 2.5× faster at $8/$40 per million tokens.
- Clearer writing: less jargon, with the important information placed at the start of its responses.
- Newer knowledge: a June 2026 reliable knowledge cutoff, up from Opus 5’s May 2026.
- Default effort is now
mediumon the API (Opus 5 defaulted tohigh), and adaptive thinking is always on.
Early customer results in Anthropic’s announcement point the same way. Deloitte reported Opus 5.5 caught 72% of known bugs in a code-review test versus 56% for Opus 5, and Hebbia measured 86.6% of its target criteria versus 60.3% for Opus 5.
Claude Opus 5.5 Specs and Pricing
| Property | Claude Opus 5.5 |
|---|---|
| API model ID | claude-opus-5-5 (Bedrock: anthropic.claude-opus-5-5) |
| Released | September 22, 2026 |
| Input / output price | $4 / $20 per million tokens |
| Prompt cache reads | $0.20 / MTok (0.05× base input) |
| Cache writes | $5 / MTok (5-minute) · $8 / MTok (1-hour) |
| Batch API | 50% off: $2 / $10 per MTok |
| Fast mode (research preview) | $8 / $40 per MTok, up to 2.5× speed — Claude API and Claude Code |
| Context window | 1M tokens (~555k words) |
| Max output | 128K tokens (300K on the Batch API with a beta header) |
| Thinking | Adaptive, always on — it can’t be turned off; steer depth with effort |
| Default effort | medium |
| Comparative latency | Moderate |
| Reliable knowledge cutoff | June 2026 |
| Data retention | Zero data retention available |
| Retirement | Not sooner than September 22, 2027 |
Model your own workload in the Claude API cost estimator, which includes Opus 5.5, its fast mode, and its cache-read rate.
Claude Opus 5.5 Benchmarks
These are Anthropic’s published launch figures. Opus 5.5 was run with adaptive thinking at max effort unless noted; its Terminal-Bench 4.0 score uses xhigh effort.
| Benchmark | Opus 5.5 | Fable 5.1 | Opus 5 | GPT-6 Astra | GPT-5.6 Sol |
|---|---|---|---|---|---|
| Terminal-Bench 4.0 (agentic coding) | 66.4% | 55.8% | 52.3% | 57.9% | 37.3% |
| FrontierCode v1.1 Main (agentic coding) | 54.4% | 50.3% | 48.0% | 53.3% | 47.5% |
| CursorBench 4.0 (agentic coding) | 57.8% | 51.8% | 46.6% | — | 41.7% |
| GDPval-AA v2.1 (knowledge work, Elo) | 1,846 | 1,735 | 1,708 | 1,542 | 1,588 |
| AutomationBench (business workflows) | 40.0% | 31.4% | 26.9% | 41.4% | 28.8% |
| Humanity’s Last Exam (with tools) | 67.7% | 65.6% | 63.6% | 57.2% | — |
| Terminal-Bench-Science 0.1 (scientific research) | 58.7% | 52.6% | 29.0% | 64.6% | 22.4% |
| OSWorld 2.1 (computer use, partial credit) | 81.8% | 80.7% | 74.0% | — | — |
| Chartography (chart recognition, with tools) | 89.0% | 88.4% | 83.4% | — | — |
Three things are worth reading into this table. First, Opus 5.5 leads Fable 5.1 on every row, at 40% of Fable’s list price. Second, it isn’t a clean sweep: OpenAI’s GPT-6 Astra is ahead on AutomationBench and on agentic scientific research. Third, Anthropic itself cautions that at this level “benchmark margins have become a less reliable guide to real-world differences,” and notes that some scores are pulled down by safeguards routing cybersecurity and biology tasks to other models. Several benchmarks here are newer versions than the ones in Fable 5.1’s September 1 table, so Fable’s numbers differ slightly from our Fable 5.1 guide.
The full head-to-head is in Claude Opus 5.5 vs Fable 5.1, and the cheaper alternative is covered in Sonnet 5.5 vs Opus 5.5.
API Changes: Migrating From Opus 5
Four changes can break code that runs on Opus 5 today. The first three also apply to Fable 5.1:
- Thinking can’t be disabled. Adaptive thinking is always on; control its depth with the
effortparameter instead. - Forced tool use returns an error.
tool_choiceof typeanyortoolis rejected — useautowith strict tool use or structured outputs. - Thinking blocks are tied to the model and the conversation, so editing earlier turns or switching models mid-conversation can invalidate them.
- The older
computer_20251124computer-use tool isn’t accepted on the Claude API and Google Cloud.
One more change alters response shape without failing requests: text Claude writes between tool calls now arrives inside thinking blocks, which are empty at the default display setting. If your app streams that text to users as progress updates, set a display value that returns it, or your UI will go quiet between tool calls.
Availability, Plans, and Claude Code
- Everywhere at launch: the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry, and Claude Platform on AWS, plus the Claude apps.
- Bigger subscription limits: alongside the launch, Anthropic raised five-hour usage limits on Pro, Max, Team, and seat-based Enterprise plans, and gave subscribers a one-time rate-limit reset they can save and use when they choose.
- Claude Code default: Opus 5.5 is the default model in Claude Code on Pro, Max, Team, Enterprise, and the Anthropic API (Claude Code v2.1.280 or later — run
claude update). Fast mode is available there too. - Zero data retention: available, unlike Fable 5.1, which carries mandatory 30-day retention.
See Is Claude AI Free? for what each plan includes, and Claude Code pricing for how usage limits work in practice.
Safety and Safeguards
- Alignment: Anthropic calls it the strongest-performing model it has tested on its automated behavioral audit, and reports it tried to circumvent containment boundaries about 85% less often than Opus 5 or Claude Mythos 5.1. External testing came from METR and other evaluators before release.
- Cybersecurity: safeguards are on by default and most cybersecurity tasks are re-routed to Claude Opus 4.8. Anthropic says an expansion of its Cyber Verification Program is coming in the following weeks.
- Biology: the same biology safeguards as Fable 5.1, with vetted labs, startups, and pharmaceutical companies able to apply for broader access through the new Life Sciences Verification Program.
- Anti-distillation: preserved-thinking safeguards block editing earlier context on API accounts created on or after August 31, 2026.
- Evaluation awareness: Anthropic notes signs that Opus 5.5 often suspects it is being evaluated, and says building evaluations that catch every failure before deployment remains unsolved.
Where Opus 5.5 Fits in the Lineup
| Dimension | Fable 5.1 | Opus 5.5 | Sonnet 5.5 | Haiku 4.5 |
|---|---|---|---|---|
| Anthropic’s positioning | Demanding reasoning, long-horizon agents | Recommended starting point | Best speed/intelligence balance | Fastest |
| Price / MTok (in / out) | $10 / $50 | $4 / $20 | $2 / $10 | $1 / $5 |
| Cache reads / MTok | $0.25 | $0.20 | $0.20 | $0.10 |
| Default effort | high | medium | high | — |
| Latency | Slower | Moderate | Fast | Fastest |
| Context / max output | 1M / 128K | 1M / 128K | 1M / 128K | 200K / 64K |
| Knowledge cutoff | Jun 2026 | Jun 2026 | Jun 2026 | Feb 2025 |
For most people the choice is now between Opus 5.5 and Sonnet 5.5, which costs half as much and lands within two points of Opus 5.5 on several of the same benchmarks. Claude Models Explained covers the full lineup, and the model selector gives a recommendation for your task.
Limitations and Caveats
- Benchmark scores were run at
maxorxhigheffort, but the API default ismedium. Expect lower scores and lower costs at the default; raise effort for hard steps. - Most cybersecurity work is routed to Opus 4.8 by default, which can surprise security teams until the expanded Cyber Verification Program arrives.
- Four breaking API changes mean migrating from Opus 5 is more than a model-ID swap.
- Fast mode is a research preview and costs double the standard rate.
Frequently Asked Questions
When was Claude Opus 5.5 released?
Anthropic released Claude Opus 5.5 on September 22, 2026. It succeeds Claude Opus 5, which moved to legacy status and remains available until at least July 24, 2027.
How much does Claude Opus 5.5 cost?
Claude Opus 5.5 costs $4 per million input tokens and $20 per million output tokens — 20% less than Opus 5’s $5/$25. Cache reads cost $0.20 per million tokens, the Batch API halves prices to $2/$10, and fast mode costs $8/$40. Anthropic estimates typical workloads run about 40% cheaper than on Opus 5 because the model also uses fewer tokens.
Is Claude Opus 5.5 better than Claude Fable 5.1?
On Anthropic’s published benchmarks, yes: Opus 5.5 leads Fable 5.1 on all nine rows, including 66.4% vs 55.8% on Terminal-Bench 4.0 and 1,846 vs 1,735 on GDPval-AA v2.1, at 60% lower list prices. Anthropic still positions Fable 5.1 for demanding reasoning and long-horizon agentic work, or when Opus 5.5 at higher effort still falls short, and cautions that benchmark margins are a less reliable guide at this level.
Is Claude Opus 5.5 the default model in Claude Code?
Yes. Claude Code defaults to Opus 5.5 on Pro, Max, Team, and Enterprise plans and on the Anthropic API. It requires Claude Code v2.1.280 or later; run claude update to upgrade. Fast mode for Opus 5.5 is also available in Claude Code.
What is Claude Opus 5.5’s default effort level?
Medium on the Claude API, down from Opus 5’s high. Adaptive thinking is always on and cannot be disabled; raise effort to high, xhigh, or max for difficult steps. Anthropic’s benchmark scores were run at max effort, or xhigh on Terminal-Bench 4.0.
What changed for developers migrating from Opus 5?
Four breaking changes: thinking cannot be disabled, forced tool use returns an error, thinking blocks are tied to the model and conversation, and the older computer_20251124 tool is not accepted on the Claude API and Google Cloud. Text between tool calls also now arrives in thinking blocks that are empty at the default display setting.
Did Anthropic change usage limits with Opus 5.5?
Yes. With the launch, Anthropic increased five-hour usage limits on Pro, Max, Team, and seat-based Enterprise plans, and gave subscribers a one-time rate-limit reset they can save and use whenever they choose.