Safety and alignment

OpenAI deploys source-to-sink controls against agent prompt injection

OpenAI describes deployed controls that constrain data exfiltration and unintended agent actions through Safe URL checks, user confirmation, blocking, and sandbox communication controls in ChatGPT products.

0 comments · 0 votesOpen discussion

Public discussion is readable by everyone. Sign in to comment, reply, or vote.

No comments yet. Start the discussion.

CURRENT ASSESSMENT · REVISION 1
AWAY FROM DOOM39confidence 60/100

Why it moved the index

Deployed source-to-sink controls constrain agent manipulation and silent data transfer even when model-level resistance fails, but efficacy is company-reported.

AUDIT TRAIL

Assessment history

  1. R1
    Away 39 · confidence 60

    New deployed agent safeguards absent from durable context.

    11 Aug 2026
SHARE THE FINDINGS

Share this page

DoomBench social sharing card for OpenAI deploys source-to-sink controls against agent prompt injection.
  1. DoomBench assesses “OpenAI deploys source-to-sink controls against agent prompt injection” as evidence moving away from doom, with magnitude 39 and confidence 60 out of 100 in the safety and alignment category.

  2. The DoomBench assessment of “OpenAI deploys source-to-sink controls against agent prompt injection” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.

  3. DoomBench summarizes “OpenAI deploys source-to-sink controls against agent prompt injection” as follows: OpenAI describes deployed controls that constrain data exfiltration and unintended agent actions through Safe URL checks, user...

    https://www.doombench.com/news/openai-deploys-source-to-sink-controls-against-agent-prompt-injection-2026-03-11