OpenAI · OpenAI o3

OpenAI o3

READER SUMMARY

OpenAI's April 2025 frontier reasoning model with multimodal analysis and agentic use of ChatGPT and API tools.

DOOM SCORE74.7out of 100model risk profile, not the overall index
CURRENT ASSESSMENT · REVISION 2

Why this model scores 74.7

Frontier reasoning and autonomous composition of web, code, file and image tools substantially expand capability and misuse potential at broad hosted scale, with provider controls retaining some leverage.

Capability90
Autonomy83
Deployment99
Misuse potential81
Control difficulty69
0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

MODEL-ATTRIBUTED EVIDENCE

News tied to OpenAI o3

The model score of 74.7 rates this model's risk profile. The overall Doom Index of 63.5 measures the recalibrated complete evidence corpus. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION-0.04

Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.

AutonomyTOWARD

OpenAI releases o3 and o4-mini with agentic tool use

OpenAI released o3 and o4-mini in ChatGPT and its APIs with visual reasoning and agentic use of web search, Python, files, image generation and other tools, alongside a system card under its revised Preparedness Framework.

Full item contribution
+0.03
OpenAI o3 equal share
+0.02
Read assessment →
AUDIT TRAIL

Model score history

  1. R2
    Doom Score 74.7

    Full-corpus evidence recalculated after run intensive-backfill-20260812-064900-june-2025-pass1 under bounded-corpus-v2.

    12 Aug 2026
  2. R1
    Doom Score 84.0

    Initial exact profile for the April 2025 o3 release.

    12 Aug 2026