← Model benchmark
Moonshot AI · Kimi

Kimi K3

READER SUMMARY

Moonshot AI's frontier Kimi model, assessed by UK and US government evaluators for exploit development and autonomous multi-step cyber operations.

DOOM SCORE80.4out of 100model risk profile, not the overall index
CURRENT ASSESSMENT · REVISION 1

Why this model scores 80.4

Kimi K3 trails the latest US closed models but completed a long simulated corporate attack once in ten runs and its safeguards did not block offensive cyber work. Public availability and planned open weights increase deployment and residual control difficulty, while limited reliability and the preliminary scope of testing constrain the capability and autonomy scores.

Capability82
Autonomy75
Deployment92
Misuse potential76
Control difficulty78
MODEL-ATTRIBUTED EVIDENCE

News tied to Kimi K3

The model score of 80.4 rates this model's risk profile. The overall Doom Index of 67.7 measures the accumulated evidence trajectory. These values answer different questions.

NET MODEL-ATTRIBUTED INDEX CONTRIBUTION+2.73

Each article's signed Doom Index movement is divided equally among the exact models named on that article. This prevents multi-model evidence from being counted in full on several model pages. The model's risk score itself never changes the overall index.

AUDIT TRAIL

Model score history

  1. R1
    Doom Score 80.4

    Initial independent source-backed profile for the exact released model.

    11 Aug 2026