← Model benchmark
Anthropic · Claude Opus

Claude Opus 4

Anthropic's strongest 2025 coding and agent model, designed for long-running workflows that can continue for hours and thousands of steps.

DOOM SCORE74.2out of 100
CURRENT ASSESSMENT · REVISION 1

Why this model scores 74.2

Sustained long-horizon work and tool use raised autonomy risk, while ASL-3 deployment controls and narrower access reduced it relative to capability alone.

Capability83
Autonomy78
Deployment76
Misuse potential66
Control difficulty62
AUDIT TRAIL

Model score history

  1. R1
    Doom Score 74.2

    Initial source-backed model assessment

    11 Aug 2026