SpaceXAI · Grok 3

Grok 3

READER SUMMARY

SpaceXAI's flagship third-generation beta model, combining extensive pretraining with reinforcement-learned test-time reasoning and improved mathematics, coding and instruction following.

DOOM SCORE74.4out of 100model risk profile, not the overall index
CURRENT ASSESSMENT · REVISION 1

Why this model scores 74.4

Large capability gains and minutes-long self-correcting reasoning raise capability and autonomy. Early user rollout makes deployment material but below generally available services. Broad X-linked access and strong coding increase misuse exposure, while hosted access limits downstream control difficulty.

Capability84
Autonomy62
Deployment79
Misuse potential75
Control difficulty63
0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

MODEL-ATTRIBUTED EVIDENCE

News tied to Grok 3

The model score of 74.4 rates this model's risk profile. The overall Doom Index of 63.2 measures the recalibrated complete evidence corpus. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.02

Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.

AUDIT TRAIL

Model score history

  1. R1
    Doom Score 74.4

    Initial exact profile for the February 2025 flagship Grok 3 beta tier.

    12 Aug 2026