Safety and alignment

Guidelight finds frontier AI controls only partially implemented

Guidelight AI Standards assessed six foundational control practices across Anthropic, OpenAI, Google, xAI, and Meta using publicly available system cards, safety frameworks, risk reports, and collaboration descriptions. No company scored above 2.50 out of 5 overall, and the assessment found particularly weak public evidence for prevention and containment practices.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
TOWARD DOOM55confidence 76/100

Why it moved the index

Toward-doom magnitude 55: an independent cross-lab review finding only partial implementation of logging, monitoring, gated actions, circuit breaking, third-party review, and containment indicates a material control gap at frontier developers. Confidence 76: the dated assessment publishes its methodology, scores, and reviewed sources, but it measures documented public evidence and cannot prove that undisclosed internal practices are absent.

AUDIT TRAIL

Assessment history

  1. R1
    Toward 55 · confidence 76

    Adds a dated, method-published cross-lab assessment of foundational AI control practices that is absent from the durable evidence corpus.

    24 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for Guidelight finds frontier AI controls only partially implemented.
  1. DoomBench assesses “Guidelight finds frontier AI controls only partially implemented” as evidence moving toward doom, with magnitude 55 and confidence 76 out of 100 in the safety and alignment category.

  2. The DoomBench assessment of “Guidelight finds frontier AI controls only partially implemented” is based on reporting from Guidelight AI Standards and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “Guidelight finds frontier AI controls only partially implemented” as follows: Guidelight AI Standards assessed six foundational control practices across Anthropic, OpenAI, Google, xAI, and Meta using publicly...

    https://www.doombench.com/news/guidelight-finds-frontier-ai-controls-only-partially-implemented-2026-08-18