GPT-5
READER SUMMARYOpenAI's broadly deployed flagship reasoning and tool-use model, used by the Aardvark security agent during this window.
Why this model scores 68.5
Frontier reasoning, coding, tool use, and universal ChatGPT distribution support high capability, autonomy, deployment, and misuse, while safeguards moderate control difficulty.
News tied to GPT-5
The model score of 68.5 rates this model's risk profile. The overall Doom Index of 61.2 measures the recalibrated complete evidence corpus. These values answer different questions.
Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.
Aardvark finds real vulnerabilities with a continuously running GPT-5 security agent
OpenAI introduced Aardvark, a GPT-5 security agent that continuously analyzes repositories, validates vulnerabilities, proposes patches, and had already produced CVEs.
- Full item contribution
- -0.19
- GPT-5 equal share
- -0.19
Model score history
- R1Doom Score 68.5
New exact model needed for the accepted Aardvark relationship and absent from the durable catalogue.
11 Aug 2026