Mistral AI ๐Ÿ‡ซ๐Ÿ‡ท ยท Shieldstral

Shieldstral 1.0 3B

READER SUMMARY

A 3B open-weight multimodal safety classifier that applies plain-language policies to prompts, responses, refusals, toxicity, and images.

DOOM SCORE20.1out of 100model risk profile, not the overall index
CURRENT ASSESSMENT ยท REVISION 1

Why this model scores 20.1

Shieldstral is highly deployable because its Apache 2.0 weights fit a single 16 GB GPU, but it is a narrow moderation classifier with no demonstrated autonomous agency or general-purpose capability.

Capability20
Autonomy1
Deployment78
Misuse potential8
Control difficulty5
0 comments ยท 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

MODEL-ATTRIBUTED EVIDENCE

News tied to Shieldstral 1.0 3B

The model score of 20.1 rates this model's risk profile. The overall Doom Index of 61.4 measures the complete temporally weighted evidence record. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION-0.15

Each article's current Doom Index contribution is divided equally among the exact models named on that article. This prevents multi-model evidence from being claimed in full on several model pages. Model risk scores use a bounded temporal offset around their technical profile, but never feed back into the overall index.

AUDIT TRAIL

Model score history

  1. R1
    Doom Score 20.1

    Mistral's dated release, official model card, and weight repository establish Shieldstral 1.0 3B as an open, policy-adaptive safety classifier.

    13 Aug 2026