Safety and alignment

OpenAI publishes the GPT-4o System Card

OpenAI documented GPT-4o preparedness evaluations, external red teaming, audio safeguards, prohibited voice-output controls, and residual risks across persuasion, biological threats, cybersecurity, and model autonomy.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM41confidence 96/100

Why it moved the index

The system card made frontier multimodal risk evidence and deployed mitigations public, improving accountability and control knowledge while also documenting important residual limitations.

AUDIT TRAIL

Assessment history

  1. R1
    Away 41 · confidence 96

    New August 2024 frontier-model safety and control disclosure; exact checkpoint intentionally not inferred.

    12 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for OpenAI publishes the GPT-4o System Card.
  1. DoomBench assesses “OpenAI publishes the GPT-4o System Card” as evidence moving away from doom, with magnitude 41 and confidence 96 out of 100 in the safety and alignment category.

  2. The DoomBench assessment of “OpenAI publishes the GPT-4o System Card” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “OpenAI publishes the GPT-4o System Card” as follows: OpenAI documented GPT-4o preparedness evaluations, external red teaming, audio safeguards, prohibited voice-output controls, and residual risks across...

    https://www.doombench.com/news/openai-publishes-the-gpt-4o-system-card-2024-08-08