Anthropic launches Claude Fable 5.1 and Mythos 5.1, cutting cache-read pricing 75% but raising net per-task cost ~20% due to higher output token usage
Anthropic released Claude Fable 5.1 and Claude Mythos 5.1 as new flagship models for coding and knowledge work. Input/output/cache-write pricing unchanged ($10/$50/$12.5 per MTok), but cache read price cut 75% (from $1.00 to $0.25 per MTok). Benchmarks improved substantially (e.g., Terminal-Bench-Science 24.7%→52.6%, HLE 55.5%→59.1%, AA Intelligence Index 66 vs Fable 5's 62), but output token usage rose ~1.7x, resulting in a net ~20% higher cost per task ($3.76 vs lower for Fable 5 max). Community analysis (unconfirmed by Anthropic) suggests Fable 5.1 and Mythos 5.1 may share identical underlying weights, differentiated only by safety classifier routing/thresholds.
Entities: Anthropic, Claude Fable 5.1, Claude Mythos 5.1, Claude Opus 5, Claude Opus 4.8, Claude Fable 5
0 primary
What happened
Anthropic released two new flagship models, Claude Fable 5.1 and Claude Mythos 5.1, for coding and knowledge work. Base input, output and cache-write pricing stayed the same ($10/$50/$12.5 per million tokens), while cache-read pricing dropped 75% (from $1.00 to $0.25 per MTok). Benchmark scores rose substantially (Terminal-Bench-Science 24.7% to 52.6%, HLE 55.5% to 59.1%), but because the models generate roughly 1.7x more output tokens per task, the net cost per task actually increased by about 20% compared with Fable 5.
Why it matters
Developers and enterprises budgeting for Claude-based agentic or long-context workflows need to recalculate costs rather than assume the cache-price cut lowers their bills, since higher output volume likely outweighs the cache savings for typical tasks. This is a useful, concrete warning against taking headline pricing cuts at face value. Competitive impact is modest: this reshuffles benchmark rankings among Anthropic, OpenAI and xAI's existing frontier models rather than shifting market power to a new player.
What is noise
The "new SOTA" and "reclaiming the frontier lead" framing glosses over the fact that overall task cost went up, not down, which undercuts the pricing announcement as a customer win. The claim that Fable 5.1 and Mythos 5.1 share identical underlying weights, differentiated only by safety classifier routing, is explicitly unconfirmed community speculation and should not be treated as fact. No primary source links (official Anthropic pricing page, benchmark methodology, or model card) were captured, so all figures trace back to a newsletter aggregating tweets and third-party Artificial Analysis measurements.
Watch next
- 01Whether Anthropic publishes an official model card or benchmark methodology confirming the AA Intelligence Index 66 figure, rather than relying on Artificial Analysis's third-party measurement
- 02Independent developer cost reports over the next 4-6 weeks testing whether the ~20% net cost increase per task holds across real agentic workloads, not just Anthropic's own benchmark tasks
- 03Whether Anthropic confirms or denies the community claim that Fable 5.1 and Mythos 5.1 share identical weights differentiated only by safety routing, which would materially change how developers choose between them
Coverage
1 storyMore capability signals
Full feed →- AI systems outperform expert humans in persuasive communication22 Jun 202681
- WIRED investigation: Flock Safety's AI person-search tools let police run broad description-based surveillance, with weak guardrails against misuse3 Sept 202680
- Hcompany open-sources NeoMME, a from-scratch multimodal-native encoder family, and NeoMME-Retriever for visual document retrieval3 Sept 202679
- Benchmark results show significant improvement in AI agent performance on WorkBench15 Jun 202679