Safety and alignment

OpenAI revises framework for severe frontier-model risks

OpenAI released Preparedness Framework version 2 with tracked catastrophic-risk categories, safeguards reports, Safety Advisory Group review and new research categories including long-range autonomy, sandbagging and safeguard undermining.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM34confidence 72/100

Why it moved the index

The revised framework makes severe-risk evaluation and safeguards review more operational, while its provider-controlled and revisable nature limits confidence in realized control strength.

AUDIT TRAIL

Assessment history

  1. R1
    Away 34 · confidence 72

    Material successor to the durable 2023 framework event with distinct categories, governance and deployment criteria.

    12 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for OpenAI revises framework for severe frontier-model risks.
  1. DoomBench assesses “OpenAI revises framework for severe frontier-model risks” as evidence moving away from doom, with magnitude 34 and confidence 72 out of 100 in the safety and alignment category.

  2. The DoomBench assessment of “OpenAI revises framework for severe frontier-model risks” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “OpenAI revises framework for severe frontier-model risks” as follows: OpenAI released Preparedness Framework version 2 with tracked catastrophic-risk categories, safeguards reports, Safety Advisory Group review and...

    https://www.doombench.com/news/openai-revises-framework-for-severe-frontier-model-risks-2025-04-15