Claude 2
READER SUMMARYA 100,000-token frontier conversational model with improved coding, mathematics and reasoning, available through API and public web beta.
Why this model scores 54.5
Long context and stronger general performance raised capability and agent usefulness. Hosted access, safety training and geographic limits retained meaningful centralized control.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
News tied to Claude 2
The model score of 54.5 rates this model's risk profile. The overall Doom Index of 61.5 measures the recalibrated complete evidence corpus. These values answer different questions.
Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.
Anthropic finds long contexts enable many-shot jailbreaking
Anthropic showed that hundreds of in-prompt demonstrations could override safety training across several large language models, disclosed the weakness to peers and deployed prompt-classification mitigations that reduced one measured attack rate from 61 percent to 2 percent.
- Full item contribution
- +0.03
- Claude 2 equal share
- +0.03
Anthropic launches Claude 2 through API and public web access
Anthropic released Claude 2 with improved reasoning and coding, a 100,000-token context window, API access and a public beta website in the United States and United Kingdom.
- Full item contribution
- +0.03
- Claude 2 equal share
- +0.03
Model score history
- R2Doom Score 54.5
Full-corpus evidence recalculated after run intensive-backfill-20260812-033703-april-2024-incremental under bounded-corpus-v2.
12 Aug 2026 - R1Doom Score 50.6
New exact Claude 2 model profile from dated launch evidence.
12 Aug 2026