dogear

enter for all results · esc to close

Continuously Hardening ChatGPT Atlas Against Prompt Injection Attacks

openai.comsite

OpenAI's Dec 2025 disclosure of a real attack chain (malicious email → agent sends resignation letter) and the RL-trained automated attacker they built to find new injection classes before external adversaries do. OpenAI explicitly states deterministic guarantees are not…

from
Prompt Injection
added
2026-10-10
likes
0

Prompt Injection › Articles and Blog posts: “OpenAI's Dec 2025 disclosure of a real attack chain (malicious email → agent sends resignation letter) and the RL-trained automated attacker they built to find new injection classes before external adversaries do. OpenAI explicitly states deterministic guarantees are not…”