Safety and alignment

OpenAI documents operational controls for autonomous Codex agents

OpenAI reported using sandboxing, approval review, managed network policies, credential controls, agent-native telemetry, and human security review in its internal Codex deployment.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM44confidence 82/100

Why it moved the index

Layered technical boundaries, approvals, network controls, credential handling, and audit trails reduce risk from operational coding agents. Evidence is from the developer's internal deployment rather than independent outcome evaluation.

AUDIT TRAIL

Assessment history

  1. R1
    Away 44 · confidence 82

    New practical evidence of layered controls used in an operational autonomous-agent deployment.

    11 Aug 2026