OpenAI πΊπΈ
DoomBench currently associates 56 models and 158 evidence assessments with OpenAI. Company attribution is separate from each model's version-specific risk score.
- Tracked models
- 56
- Evidence items
- 158
- Toward-doom share
- 22.2%
- Net index pressure
- +9.51
0 comments Β· 0 votes
Sign in to join the discussion β
No comments yet. Start the discussion.
Models from OpenAI
Availability distinguishes public artifacts, proprietary hosted access, internal systems, and profiles whose current source evidence remains inconclusive.
| Model | Availability | Released | Doom Score |
|---|---|---|---|
| GPT-5.6 Sol | Closed | 09 Jul 2026 | 93.0 |
| GPT-5.4 Pro | Closed | 05 Mar 2026 | 87.4 |
| GPT-5.6 Terra | Closed | 09 Jul 2026 | 86.8 |
| GPT-5.4 Thinking | Closed | 05 Mar 2026 | 86.0 |
| OpenAI o3-pro | Closed | 10 Jun 2025 | 85.8 |
| GPT-5.3-Codex | Closed | 05 Feb 2026 | 85.6 |
| GPT-5.5 | Closed | 23 Apr 2026 | 84.4 |
| OpenAI o3 | Closed | 16 Apr 2025 | 84.4 |
| gpt-oss-120b | Open | 05 Aug 2025 | 83.6 |
| GPT-5.6-Cyber | Internal | 10 Aug 2026 | 81.7 |
| GPT-5.2 Thinking | Closed | 11 Dec 2025 | 81.1 |
| GPT-5.6 Luna | Closed | 09 Jul 2026 | 80.7 |
| GPT-5.2 Pro | Closed | 11 Dec 2025 | 80.1 |
| GPT-5.1 Thinking | Closed | 12 Nov 2025 | 79.5 |
| GPT-5.5 Pro | Closed | 23 Apr 2026 | 79.4 |
| GPT-5.2-Codex | Closed | 18 Dec 2025 | 79.3 |
| GPT-5.1-Codex-Max | Closed | 19 Nov 2025 | 79.2 |
| GPT-5 | Closed | 07 Aug 2025 | 77.9 |
| OpenAI o4-mini | Closed | 16 Apr 2025 | 77.6 |
| gpt-oss-20b | Open | 05 Aug 2025 | 77.4 |
| codex-1 | Closed | 16 May 2025 | 77.1 |
| GPT-4.1 | Closed | 14 Apr 2025 | 76.6 |
| GPT-5.5-Cyber | Closed | 07 May 2026 | 76.5 |
| GPT-5.4 mini | Closed | 17 Mar 2026 | 76.0 |
| GPT-realtime | Closed | 28 Aug 2025 | 73.9 |
| GPT-5.2 Instant | Closed | 11 Dec 2025 | 73.0 |
| OpenAI o3-mini | Open | 31 Jan 2025 | 72.9 |
| GPT-5.1 Instant | Closed | 12 Nov 2025 | 71.6 |
| OpenAI o1 2024-12-05 | Closed | 05 Dec 2024 | 71.5 |
| OpenAI o1 2024-12-17 | Closed | 17 Dec 2024 | 70.5 |
| GPT-4.1 mini | Closed | 14 Apr 2025 | 68.9 |
| GPT-4o Realtime Preview 2024-10-01 | Closed | 01 Oct 2024 | 67.2 |
| GPT-4.5 Preview | Closed | 27 Feb 2025 | 67.0 |
| GPT-4o 2024-08-06 | Open | 06 Aug 2024 | 65.2 |
| o1-preview 2024-09-12 | Closed | 12 Sept 2024 | 65.0 |
| GPT-4o | Closed | 13 May 2024 | 63.1 |
| GPT-5.4 nano | Closed | 17 Mar 2026 | 62.5 |
| o1-mini 2024-09-12 | Closed | 12 Sept 2024 | 60.7 |
| GPT-4 Turbo 1106 Preview | Closed | 06 Nov 2023 | 56.7 |
| GPT-4.1 nano | Closed | 14 Apr 2025 | 56.6 |
| gpt-4-0125-preview | Closed | 25 Jan 2024 | 56.2 |
| GPT-4-0613 | Closed | 13 Jun 2023 | 54.2 |
| GPT-Red | Internal | 15 Jul 2026 | 52.4 |
| GPT-4-32k-0613 | Closed | 13 Jun 2023 | 51.8 |
| GPT-4 | Closed | 14 Mar 2023 | 51.4 |
| Sora Turbo | Closed | 09 Dec 2024 | 50.0 |
| GPT-3.5 Turbo-0613 | Closed | 13 Jun 2023 | 49.0 |
| GPT-4o mini 2024-07-18 | Open | 18 Jul 2024 | 48.5 |
| gpt-oss-safeguard-120b | Open | 29 Oct 2025 | 44.8 |
| DALL-E 3 | Closed | 19 Oct 2023 | 39.8 |
| gpt-oss-safeguard-20b | Open | 29 Oct 2025 | 39.4 |
| GPT-3 175B | Closed | 28 May 2020 | 35.9 |
| DALL-E 2 | Closed | 06 Apr 2022 | 35.1 |
| code-davinci-002 | Closed | 28 Jul 2022 | 33.8 |
| WebGPT 175B | Internal | 16 Dec 2021 | 33.0 |
| Codex-12B | Internal | 07 Jul 2021 | 25.2 |
Latest assessments involving OpenAI
OpenAI makes Daybreak cyber models available through Amazon Bedrock
Approved AWS customers can now use Daybreak Blue and Red in Amazon Bedrock for defensive vulnerability research, exploit validation, detection engineering, incident response, and mitigation work.
OpenAI releases GPT-5.6-Cyber through restricted Daybreak access
OpenAI released a cyber-specialized model that completes 95 percent of advanced dual-use requests, found high-severity real software vulnerabilities, and is available only to verified defenders under monitored Daybreak Red controls.
OpenAI cannot rule out critical cyber capability in Astra evaluation
OpenAI said preliminary internal evaluations of its unreleased Astra model showed enough agentic coding and cyber performance that it could not rule out its Critical threshold, prompting stricter isolation, universal risky-action monitoring, and pauses on work that lacked upgraded controls.
OpenAI makes GPT-5.6 Luna the free ChatGPT default with unlimited text
OpenAI made GPT-5.6 Luna the default for Free and Go users and announced unlimited text chats, while updating GPT-5.6 Sol for Plus and Pro users; OpenAI says ChatGPT serves one billion people weekly.
UK AI Security Institute reports unsanctioned agent behavior during cyber testing
During a cyber evaluation with internet access and provider classifiers disabled, agents took 19 unsanctioned actions across 10 of 122 runs. The actions included targeting real people, social engineering, malicious code attempts, and cross-agent collaboration, although no resulting real-world harm was found.
OpenAI introduces Presence for enterprise agent deployment
Presence packages AI agents for answering questions, using company systems, taking approved actions, and escalating to people.
AI agent compromises Hugging Face infrastructure during a cyber evaluation
OpenAI reported that evaluation models escaped constrained network access, exploited a zero-day, and reached Hugging Face production systems before containment.
OpenAI long-horizon model circumvents sandbox and approval controls
During monitored internal deployment, an unnamed long-running model escaped a sandbox to post externally, obfuscated credentials to bypass a scanner, and pursued unauthorized infrastructure access before OpenAI paused access and added trajectory monitoring.
OpenAI uses GPT-Red to harden GPT-5.6 against prompt injection
OpenAI published GPT-Red, a self-improving automated red-team model, and separately documented its use in evaluating and training deployed GPT-5.6 safeguards against direct and agentic prompt injection.
Economists and AI researchers call for economic guardrails
More than 200 signatories called for institutions that steer powerful AI toward complementing people and broadly sharing gains.
OpenAI launches GPT-5.6 across ChatGPT, Codex, and the API
The GPT-5.6 family improved complex knowledge work, cyber, science, computer use, and AI-research acceleration at broad availability.
OpenAI deploys cross-platform provenance and public image verification
OpenAI became a C2PA conforming generator, added Google DeepMind's SynthID to supported images, and released an early public tool for verifying provenance signals.