OpenAI deploys source-to-sink controls against agent prompt injection
OpenAI describes deployed controls that constrain data exfiltration and unintended agent actions through Safe URL checks, user confirmation, blocking, and sandbox communication controls in ChatGPT products.
0 comments · 0 votes
Sign in to join the discussion →
No comments yet. Start the discussion.
Why it moved the index
Deployed source-to-sink controls constrain agent manipulation and silent data transfer even when model-level resistance fails, but efficacy is company-reported.
Assessment history
-
R1
Away 39 · confidence 60
New deployed agent safeguards absent from durable context.
11 Aug 2026
Share this page
-
DoomBench assesses “OpenAI deploys source-to-sink controls against agent prompt injection” as evidence moving away from doom, with magnitude 39 and confidence 60 out of 100 in the safety and alignment category.
-
The DoomBench assessment of “OpenAI deploys source-to-sink controls against agent prompt injection” is based on reporting from OpenAI and records the editorial rationale, source quality, attribution, and revision history.
-
DoomBench summarizes “OpenAI deploys source-to-sink controls against agent prompt injection” as follows: OpenAI describes deployed controls that constrain data exfiltration and unintended agent actions through Safe URL checks, user...
https://www.doombench.com/news/openai-deploys-source-to-sink-controls-against-agent-prompt-injection-2026-03-11