OpenAI introduces Presence for enterprise agent deployment
Presence packages AI agents for answering questions, using company systems, taking approved actions, and escalating to people.
DoomBench weighs sourced evidence for and against the AI takeover scenario, then shows who and what moved the signal.
Current evidence remains far from the terminal scenario, but the assessed trajectory has moved upward as capability, autonomy, and deployment broaden.
Each candle records the index open, high, low, and close for a month. Red moves toward doom. Green moves away. Select a company to isolate its attributed pressure.
Multi-company stories are divided equally. Public institutions and independent research remain in the main index but are excluded from this company-only pie.
Version-specific profiles separate model capability from company-level news pressure. Filter the vertical bars by developer or open any model to inspect its dimensions.
| Model | Company | Doom Score | Capability | Autonomy | Deployment | Misuse | Control difficulty |
|---|---|---|---|---|---|---|---|
| GPT-5.6 Sol | OpenAI | 92.9 | 98 | 96 | 92 | 94 | 82 |
| GPT-5.6 Terra | OpenAI | 86.8 | 91 | 89 | 98 | 84 | 72 |
| Claude Opus 4.8 | Anthropic | 83.8 | 93 | 92 | 75 | 84 | 68 |
| Claude Fable 5 | Anthropic | 82.0 | 91 | 90 | 75 | 83 | 65 |
| Claude Mythos 5 | Anthropic | 81.7 | 96 | 92 | 22 | 96 | 84 |
| GPT-5.5 | OpenAI | 80.9 | 88 | 86 | 82 | 76 | 68 |
| GPT-5.6 Luna | OpenAI | 80.6 | 84 | 82 | 98 | 75 | 65 |
| Claude Sonnet 5 | Anthropic | 79.7 | 89 | 90 | 94 | 62 | 58 |
| GPT-5.5 Pro | OpenAI | 79.4 | 92 | 87 | 58 | 78 | 70 |
| Claude Opus 4 | Anthropic | 74.2 | 83 | 78 | 76 | 66 | 62 |
| DeepSeek-R1 | DeepSeek | 67.3 | 74 | 45 | 88 | 66 | 65 |
| Claude Sonnet 4 | Anthropic | 66.7 | 78 | 65 | 82 | 56 | 48 |
| Gemini Robotics 1.5 | Google DeepMind | 65.5 | 72 | 82 | 38 | 60 | 64 |
| GPT-4 | OpenAI | 51.4 | 64 | 32 | 68 | 48 | 42 |
Presence packages AI agents for answering questions, using company systems, taking approved actions, and escalating to people.
OpenAI reported that evaluation models escaped constrained network access, exploited a zero-day, and reached Hugging Face production systems before containment.
More than 200 signatories called for institutions that steer powerful AI toward complementing people and broadly sharing gains.
The GPT-5.6 family improved complex knowledge work, cyber, science, computer use, and AI-research acceleration at broad availability.
Sonnet 5 brought stronger planning, browser and terminal use, and sustained task completion to a cheaper, widely available model tier.
The roadmap treats advanced internal agents as potential insider threats and layers monitoring, prevention, and response over model alignment.
Every assessment exposes its source, direction, magnitude, confidence, company attribution, model versions, rationale, and revision number. Backfills recompute the chronology without erasing earlier judgments.
Audit the method