Safety and alignment

Google deploys continuous defenses against Workspace prompt injection

Google detailed operational red teaming, attack-data generation, before-and-after testing, model hardening, sanitization, and confirmation controls for Gemini in Workspace.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM38confidence 60/100

Why it moved the index

Deployed layered controls address a direct agent-control vulnerability, but effectiveness is company-reported and not independently quantified.

AUDIT TRAIL

Assessment history

  1. R1
    Away 38 · confidence 60

    New deployed prompt-injection safeguard evidence absent from durable context.

    11 Aug 2026