OpenAI πŸ‡ΊπŸ‡Έ Β· GPT

GPT-5 mini

A smaller hosted GPT-5 reasoning model released for lower-cost, lower-latency API workloads with tool calling and agentic-task support. NIST used it as one of four U.S. reference models in its DeepSeek evaluation.

DOOM SCORE73.0out of 100model risk profile, not the overall index
0 comments Β· 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT Β· REVISION 1

Why this model scores 73.0

OpenAI's dated launch page documents strong reasoning and coding benchmarks, API availability, broad built-in tools, parallel tool calling, and low pricing. It remains a proprietary hosted model; NIST's evaluation provides comparative security evidence but does not establish a real-world compromise.

Capability78
Autonomy70
Deployment100
Misuse potential68
Control difficulty52
MODEL-ATTRIBUTED EVIDENCE

News tied to GPT-5 mini

The model score of 73.0 rates this model's risk profile. The overall Doom Index of 61.3 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.04

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

Safety TOWARD

NIST finds DeepSeek agents highly vulnerable to simulated hijacking

NIST's CAISI evaluated three DeepSeek models and four U.S. reference models on 19 benchmarks. In controlled AgentDojo simulations, agents using DeepSeek-R1-0528 were 12 times more likely than GPT-5 and Claude Opus 4 agents to follow malicious instructions, while the model complied with 94% of jailbreak requests versus 8% for U.S. references. The tests did not document a real-world escape or compromise.

Full item contribution
+0.24
GPT-5 mini equal share
+0.04
Read assessment β†’
AUDIT TRAIL

Model score history

  1. R1
    Doom Score 73.0

    New exact-version profile required by the NIST comparison, supported by OpenAI's dated release and hosted API availability evidence.

    14 Aug 2026