Safety and alignment

Turing report says model-only evaluations miss deployed AI control risks

The Alan Turing Institute published a frontier-risk report arguing that evaluating a model in isolation cannot verify the behavior of complete deployed AI systems. It calls for assurance of technology, people, processes, monitorability, and safe fallback arrangements, while stressing that a proposed kill switch needs proof and governance before it can be trusted.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
TOWARD DOOM24confidence 68/100

Why it moved the index

The report identifies a direct human-control gap: model-only tests do not establish that autonomous systems remain monitorable, governable, and safely overridable after deployment. This is a specific institutional synthesis and research agenda, not proof that its proposed assurance methods already work or that another system escaped.

AUDIT TRAIL

Assessment history

  1. R1
    Toward 24 · confidence 68

    New, dated whole-system control analysis distinct from the existing Turing funding announcement and recorded agent incidents.

    25 Sept 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Turing report says model-only evaluations miss deployed AI control risks.
  1. DoomBench assesses “Turing report says model-only evaluations miss deployed AI control risks” as evidence moving toward doom, with magnitude 24 and confidence 68 out of 100 in the safety and alignment category.

  2. The DoomBench assessment of “Turing report says model-only evaluations miss deployed AI control risks” is based on reporting from The Alan Turing Institute and records the editorial rationale, source quality, attribution, and revision...

  3. DoomBench summarizes “Turing report says model-only evaluations miss deployed AI control risks” as follows: The Alan Turing Institute published a frontier-risk report arguing that evaluating a model in isolation cannot verify the...

    https://www.doombench.com/news/turing-report-says-model-only-evaluations-miss-deployed-ai-control-risks-2026-09-23