Anthropic · Claude Mythos

Claude Mythos Preview

READER SUMMARY

A restricted general-purpose frontier preview with exceptional autonomous vulnerability discovery and exploitation capability, deployed to vetted Project Glasswing cyber defenders.

DOOM SCORE77.7out of 100model risk profile, not the overall index
CURRENT ASSESSMENT · REVISION 2

Why this model scores 77.7

Claude Mythos Preview surpassed nearly all human experts on key vulnerability tasks and completed sustained simulated network attacks, while live restricted deployment found thousands of severe flaws. Strict partner access keeps deployment low, but the capability's offensive dual use and acknowledged lack of robust general-release safeguards create high misuse and residual control difficulty.

Capability94
Autonomy88
Deployment15
Misuse potential98
Control difficulty84
MODEL-ATTRIBUTED EVIDENCE

News tied to Claude Mythos Preview

The model score of 77.7 rates this model's risk profile. The overall Doom Index of 62.1 measures the recalibrated complete evidence corpus. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.53

Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.

AUDIT TRAIL

Model score history

  1. R2
    Doom Score 77.7

    Full-corpus evidence recalculated after run intensive-backfill-20260811-191225 under bounded-corpus-v2.

    11 Aug 2026
  2. R1
    Doom Score 79.6

    Initial exact-tier profile from the dated model assessment and separately verified Project Glasswing deployment results.

    11 Aug 2026