Anthropic launches Claude Opus 5.5, matching Fable 5.1 performance at ~40% lower operating cost with faster output and less "Claudish" writing
Anthropic released Claude Opus 5.5, available now via Claude Platform (model ID claude-opus-5-5), AWS, Google Cloud, and Microsoft Azure. It cuts token prices 20% ($4/$20 per million input/output tokens vs $5/$25 for Opus 5) and cache read costs 60%, uses fewer tokens and generates output 30%+ faster, increases five-hour usage limits by 20% for subscribers, adds safety safeguards (cybersecurity/biology/frontier-LLM routing to other models), introduces "Preserved Thinking" anti-distillation measures and EU AI Act-compliant watermarking, and removes the option to disable "Thinking" mode.
Entities: Anthropic, Claude Opus 5.5, Claude Fable 5.1, Claude Opus 5, Claude Sonnet 5.5, Claude Haiku 5.5
0 primary
What happened
Anthropic released Claude Opus 5.5 (model ID claude-opus-5-5), available now on the Claude Platform, AWS, Google Cloud and Microsoft Azure. It cuts token prices 20% to $4/$20 per million input/output tokens (from $5/$25 for Opus 5), cuts cache-read costs 60%, raises five-hour usage limits 20% for subscribers, and Anthropic claims it matches Fable 5.1 benchmark performance while using fewer tokens and generating output over 30% faster. The release also adds safety routing for cybersecurity, biology and frontier-LLM queries to other models, introduces anti-distillation "Preserved Thinking" measures and EU AI Act watermarking, and removes the option to disable Thinking mode.
Why it matters
This is directly actionable for anyone budgeting API spend or picking a model: lower per-token pricing plus reduced token usage compounds into a real cost cut for high-volume users, and the higher usage caps matter for subscribers hitting rate limits. Enterprises and developers building on Opus should re-run their own cost and latency comparisons rather than accept Anthropic's figures, since the pricing pressure also signals a broader industry trend of frontier labs competing on cost as much as capability. The removal of the "disable Thinking" toggle is a real workflow change for anyone who relied on non-reasoning mode for speed or simpler outputs, and this is not being flagged prominently in coverage.
What is noise
The "40% lower operating cost" figure and the "less Claudish" writing claim are self-reported and subjective, not independently verified in the sourcing available here. The headline framing of "matching Fable 5.1 performance" implies parity, but the underlying benchmark comparison, its margin of error, and which specific tasks were tested are not detailed in this evidence, so treat it as a vendor claim pending third-party confirmation.
Watch next
- 01Independent benchmark confirmation: check whether Artificial Analysis or another third party verifies the claimed ~40% lower operating cost and 30%+ faster output, not just Anthropic's own numbers
- 02OpenAI and Chinese lab pricing response over the next 4-8 weeks, since Opus 5.5's $4/$20 pricing only matters if it holds up against GPT-6 Astra/GPT-5.6 Sol and cheaper alternatives
- 03Developer reaction to the removal of the 'disable Thinking' option and the new safety-routing for cybersecurity/biology/frontier queries, which could affect latency, cost or output for specific workloads not captured in headline benchmarks
Coverage
1 storyMore capability signals
Full feed →- Deepseek releases V4.1-Flash, an open-source model that sharply cuts KV cache memory and input-processing compute for AI agents10 Sept 202682
- Anthropic threat report: Claude abused for malware, drone/missile software, mass surveillance, and industrial-scale distillation by Chinese AI labs11 Sept 202680
- WIRED investigation: Flock Safety's AI person-search tools let police run broad description-based surveillance, with weak guardrails against misuse3 Sept 202680
- Google DeepMind launches AlphaGenome Atlas, a free public database of predicted effects for 9 billion possible human genome variants8 Sept 202679