SpaceXAI · Grok 3

Grok 3 mini

READER SUMMARY

A cost-efficient Grok 3 beta reasoning tier optimized for mathematics, coding and other STEM tasks through reinforcement-learned test-time compute.

DOOM SCORE71.2out of 100model risk profile, not the overall index
CURRENT ASSESSMENT · REVISION 1

Why this model scores 71.2

The mini tier trails flagship world knowledge but demonstrates strong reasoning and coding with cost-efficient inference. Test-time reasoning raises autonomy, while early hosted rollout constrains deployment and control difficulty compared with open or fully general-availability systems.

Capability77
Autonomy64
Deployment78
Misuse potential70
Control difficulty58
0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

MODEL-ATTRIBUTED EVIDENCE

News tied to Grok 3 mini

The model score of 71.2 rates this model's risk profile. The overall Doom Index of 63.2 measures the recalibrated complete evidence corpus. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+0.02

Each article's leave-one-out Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores are recalculated from their technical profile and related news, but never feed back into the overall index.

AUDIT TRAIL

Model score history

  1. R1
    Doom Score 71.2

    Initial exact profile for the separately named February 2025 Grok 3 mini beta tier.

    12 Aug 2026