Safety and alignment

OpenAI deploys source-to-sink controls against agent prompt injection

OpenAI describes deployed controls that constrain data exfiltration and unintended agent actions through Safe URL checks, user confirmation, blocking, and sandbox communication controls in ChatGPT products.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM39confidence 60/100

Why it moved the index

Deployed source-to-sink controls constrain agent manipulation and silent data transfer even when model-level resistance fails, but efficacy is company-reported.

AUDIT TRAIL

Assessment history

  1. R1
    Away 39 · confidence 60

    New deployed agent safeguards absent from durable context.

    11 Aug 2026