Autonomy and agency

Anthropic releases Claude Sonnet 4.5 with longer autonomous task performance

Anthropic released Claude Sonnet 4.5 across its API and products, reporting state-of-the-art coding and computer-use results plus sustained focus for more than 30 hours on complex multi-step tasks. The accompanying system card places the model at the 2-to-8-hour software-engineering threshold while documenting ASL-3 safeguards and controlled alignment evaluations.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
TOWARD DOOM60confidence 88/100

Why it moved the index

A broadly deployed frontier model that can sustain complex work for more than 30 hours and improves coding, tool use, and computer operation materially raises consequential agent autonomy. Anthropic's system card also documents meaningful ASL-3 controls, improved alignment, and results below ASL-4 thresholds. Reward-hacking, sabotage, sandbagging, and related behaviors were tested in controlled evaluations; this item does not describe a real-world escape or compromise.

AUDIT TRAIL

Assessment history

  1. R1
    Toward 60 · confidence 88

    New exact-version release and system-card evidence absent from the refreshed durable context.

    14 Aug 2026