Safety and alignment

OpenAI publishes the GPT-4o System Card

OpenAI documented GPT-4o preparedness evaluations, external red teaming, audio safeguards, prohibited voice-output controls, and residual risks across persuasion, biological threats, cybersecurity, and model autonomy.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM41confidence 96/100

Why it moved the index

The system card made frontier multimodal risk evidence and deployed mitigations public, improving accountability and control knowledge while also documenting important residual limitations.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

AUDIT TRAIL

Assessment history

  1. R1
    Away 41 · confidence 96

    New August 2024 frontier-model safety and control disclosure; exact checkpoint intentionally not inferred.

    12 Aug 2026