Anthropic releases Claude Opus 5.5 with system card detailing RSP safety evaluations
Anthropic released Claude Opus 5.5 along with its system card, detailing RSP (Responsible Scaling Policy) evaluations, safety classifiers, biological/cyber risk evaluations, AI R&D autonomy assessments, and alignment/agentic safety testing. Notably, Anthropic changed policy to stop testing helpful-only model versions for bio evals, instead using refusal-avoidant test design. Evaluated capabilities are reported as on-trend improvements over the prior model, not crossing new RSP risk thresholds (still CB-1, not CB-2; still Autonomy-1, not Autonomy-2).
Entities: Anthropic, Claude Opus 5.5, Claude Opus 5, Claude Fable 5.1, Claude Mythos 5.1, Artificial Analysis
0 primary
What happened
Anthropic released Claude Opus 5.5 alongside a system card covering its Responsible Scaling Policy (RSP) evaluations: safety classifiers, bio/cyber risk tests, AI R&D autonomy checks and agentic safety testing. The card reports on-trend capability gains, not a threshold crossing: the model stays at CB-1 (not CB-2) and Autonomy-1 (not Autonomy-2). One real policy change is buried in the detail: Anthropic has stopped testing helpful-only model versions for bio risk and switched to refusal-avoidant test design instead.
Why it matters
This is a routine capability update dressed up as news, useful mainly to people who track RSP compliance and model safety testing methodology rather than general audiences. The bio-eval testing change is the one item worth attention: it alters how Anthropic checks whether a model can be coaxed into providing dangerous biological information, which matters to biosecurity researchers, red-teamers and anyone relying on Anthropic's safety claims. For developers and enterprises, the practical takeaway is a marginally better, cheaper model with no new risk classification, so no urgent action is required beyond normal model-swap evaluation.
What is noise
The "world's most powerful model" framing rests on selected benchmarks (Artificial Analysis) and a comparison against a rival model referred to inconsistently as "Fable 5.1" and "Mythos 5.1" in the source material, which suggests the underlying commentary itself has extraction or naming errors and should not be taken as a precise comparison. There are no primary links in this extraction: the analysis is a LessWrong write-up of the card, not the card itself, so specific numbers and claims have not been independently verified here. Calling this a safety milestone is overstated; the card itself says capabilities are on-trend and no risk threshold was crossed.
Watch next
- 01Read the actual Anthropic system card (not secondary commentary) to verify the bio-eval methodology change and confirm CB-1/Autonomy-1 classifications with primary source numbers
- 02Check whether independent evaluators (e.g. CAISI, UK AISI) publish their own assessment of Opus 5.5 that confirms or contradicts Anthropic's self-reported RSP results
- 03Watch for the next model release (from Anthropic or the unnamed rival) to see if the bio-eval testing change becomes an industry norm or stays a one-off Anthropic policy shift
Coverage
1 storyMore capability signals
Full feed →- Deepseek releases V4.1-Flash, an open-source model that sharply cuts KV cache memory and input-processing compute for AI agents10 Sept 202682
- OpenAI launches GPT-6 Sol and Luna at half the token price of GPT-5.6, with roughly flat intelligence scores per independent analysis22 Sept 202680
- Anthropic threat report: Claude abused for malware, drone/missile software, mass surveillance, and industrial-scale distillation by Chinese AI labs11 Sept 202680
- WIRED investigation: Flock Safety's AI person-search tools let police run broad description-based surveillance, with weak guardrails against misuse3 Sept 202680