OpenAI · GPT-5.4

GPT-5.4 Thinking

READER SUMMARY

Frontier reasoning and computer-use model deployed across ChatGPT and Codex and used as OpenAI's internal coding-agent monitor.

DOOM SCORE74.1out of 100model risk profile, not the overall index
CURRENT ASSESSMENT · REVISION 2

Why this model scores 74.1

High cyber classification, native computer use, long-horizon tools, and broad deployment create substantial risk, moderated by hosted controls and low reported chain-of-thought obfuscation.

Capability93
Autonomy90
Deployment96
Misuse potential86
Control difficulty64
MODEL-ATTRIBUTED EVIDENCE

News tied to GPT-5.4 Thinking

The model score of 74.1 rates this model's risk profile. The overall Doom Index of 62.1 measures the recalibrated complete evidence corpus. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION-0.20

Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.

AUDIT TRAIL

Model score history

  1. R2
    Doom Score 74.1

    Full-corpus evidence recalculated after run intensive-backfill-20260811-191225 under bounded-corpus-v2.

    11 Aug 2026
  2. R1
    Doom Score 86.0

    Exact missing model required for the deployed monitoring association and verified from its dated primary launch.

    11 Aug 2026