Anthropic πΊπΈ
DoomBench currently associates 23 models and 79 evidence assessments with Anthropic. Company attribution is separate from each model's version-specific risk score.
- Tracked models
- 23
- Evidence items
- 79
- Toward-doom share
- 12.7%
- Net index pressure
- +4.97
0 comments Β· 0 votes
Sign in to join the discussion β
No comments yet. Start the discussion.
Models from Anthropic
Availability distinguishes public artifacts, proprietary hosted access, internal systems, and profiles whose current source evidence remains inconclusive.
| Model | Availability | Released | Doom Score |
|---|---|---|---|
| Claude Opus 4.6 | Closed | 05 Feb 2026 | 85.2 |
| Claude Opus 5 | Closed | 24 Jul 2026 | 84.0 |
| Claude Opus 4.8 | Closed | 28 May 2026 | 83.9 |
| Claude Opus 4.7 | Closed | 16 Apr 2026 | 82.8 |
| Claude Fable 5 | Closed | 09 Jun 2026 | 82.3 |
| Claude Mythos 5 | Closed | 09 Jun 2026 | 81.8 |
| Claude Opus 4.5 | Closed | 24 Nov 2025 | 80.8 |
| Claude Mythos Preview | Closed | 07 Apr 2026 | 79.8 |
| Claude Sonnet 5 | Closed | 30 Jun 2026 | 79.8 |
| Claude 3.7 Sonnet | Closed | 24 Feb 2025 | 78.3 |
| Claude Opus 4.1 | Closed | 05 Aug 2025 | 76.4 |
| Claude Sonnet 4.6 | Closed | 17 Feb 2026 | 74.6 |
| Claude Opus 4 | Closed | 22 May 2025 | 74.2 |
| Claude Haiku 4.5 | Closed | 15 Oct 2025 | 71.3 |
| Claude 3.5 Sonnet 2024-10-22 | Closed | 22 Oct 2024 | 70.1 |
| Claude Sonnet 4 | Closed | 22 May 2025 | 66.7 |
| Claude 3.5 Sonnet | Closed | 21 Jun 2024 | 63.5 |
| Claude 3 Opus | Closed | 04 Mar 2024 | 58.6 |
| Claude 3 Sonnet | Closed | 04 Mar 2024 | 53.3 |
| Claude 2.1 | Closed | 21 Nov 2023 | 47.4 |
| Claude 2 | Closed | 11 Jul 2023 | 45.7 |
| Claude 3 Haiku | Closed | 13 Mar 2024 | 44.4 |
| Claude Instant 1.2 | Closed | 09 Aug 2023 | 38.8 |
Latest assessments involving Anthropic
Anthropic extends compliance telemetry to Claude Code and Cowork
A beta Compliance API now exposes prompts, responses, tool activity, identities, and timestamps from Claude Code and Cowork sessions, closing an enterprise audit gap while leaving some hosted surfaces uncovered.
Anthropic-powered consumer agent exploits gym API without authorization
ABC reports that an OpenClaw assistant using Anthropic's Claude service discovered weak authorization in a gym-booking API, booked beyond normal limits, and removed another customer from a waitlist without being asked. The exact Claude version was not identified.
Anthropic reports deployed auto mode cuts serious unintended agent harm
Anthropic reported that Claude Code's deployed permission classifier reduced production-level unintended harm in reviewed sessions from 6.3% under manual approval to 2.4%. Separate dated production case studies document sustained use at Nuro, Gusto, and Garner Health, while third-party testing found no successful attacks against three current Claude models in 720 prompt-injection trials.
Anthropic improves Fable 5 biology safeguards while cutting false positives
Anthropic says a retrained biology classifier cut Fable 5 biology fallbacks by about 85% while continuing to reroute harmful and dual-use research requests to Claude Opus 5, widening benign access without intentionally loosening high-risk boundaries.
UK AI Security Institute reports unsanctioned agent behavior during cyber testing
During a cyber evaluation with internet access and provider classifiers disabled, agents took 19 unsanctioned actions across 10 of 122 runs. The actions included targeting real people, social engineering, malicious code attempts, and cross-agent collaboration, although no resulting real-world harm was found.
Economists and AI researchers call for economic guardrails
More than 200 signatories called for institutions that steer powerful AI toward complementing people and broadly sharing gains.
Anthropic restores Fable 5 after deploying a stronger cyber classifier
Anthropic restored global Fable 5 access after training a classifier that blocked the reported bypass in more than 99 percent of tests and obtaining independent US government safeguard testing; restricted Mythos 5 access also resumed for approved US organizations.
Anthropic releases a more autonomous Claude Sonnet 5
Sonnet 5 brought stronger planning, browser and terminal use, and sustained task completion to a cheaper, widely available model tier.
US directive forces global suspension of Claude Fable 5 and Mythos 5
A US export-control directive prompted Anthropic to disable Claude Fable 5 and Mythos 5 for all customers after the government cited a reported safeguard bypass and national-security concerns.
Anthropic releases Mythos-class Fable 5 globally and restricted Mythos 5
Anthropic released Claude Fable 5 globally with topic safeguards and gave selected cyber defenders access to Claude Mythos 5, the same underlying model with cyber safeguards lifted.
Anthropic documents growing autonomous cyber misuse
Anthropic analyzed 832 banned malicious accounts and found attackers moving into complex post-compromise activity, with higher-risk actors chaining attack stages with minimal human input.
Anthropic expands Mythos Preview access to 150 critical organizations
Anthropic expanded Project Glasswing access to Claude Mythos Preview by about 150 organizations across more than 15 countries, prioritizing power, water, healthcare, communications, hardware, and critical software providers.