OpenAI revises framework for severe frontier-model risks
OpenAI released Preparedness Framework version 2 with tracked catastrophic-risk categories, safeguards reports, Safety Advisory Group review and new research categories including long-range autonomy, sandbagging and safeguard undermining.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why it moved the index
The revised framework makes severe-risk evaluation and safeguards review more operational, while its provider-controlled and revisable nature limits confidence in realized control strength.
Assessment history
-
R1
Away 34 · confidence 72
Material successor to the durable 2023 framework event with distinct categories, governance and deployment criteria.
12 Aug 2026
Share this page
-
DoomBench assesses “OpenAI revises framework for severe frontier-model risks” as evidence moving away from doom, with magnitude 34 and confidence 72 out of 100 in the safety and alignment category.
-
The DoomBench assessment of “OpenAI revises framework for severe frontier-model risks” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “OpenAI revises framework for severe frontier-model risks” as follows: OpenAI released Preparedness Framework version 2 with tracked catastrophic-risk categories, safeguards reports, Safety Advisory Group review and...
https://www.doombench.com/news/openai-revises-framework-for-severe-frontier-model-risks-2025-04-15