OpenAI disclosed a prompt-injection technique in which malicious instructions can self-replicate through AI-agent workflows, analogous to computer worms. The attack aims both to make a model perform an unauthorized action and to cause it to copy the malicious prompt into a public or externally accessible output channel, including outbound email, files, code comments, or reposted messages, where it can reach additional agents or users.
OpenAI demonstrated email- and filesystem-based propagation with internal GPT-5.4-mini research checkpoints, plus a multi-hop scenario involving GPT-5.5 in a Codex harness with Slack-connected actions. The company said all demonstrations used simulated tool calls in training and evaluation environments, with no observed external impact or real-world incident. It has added self-reproduction objectives to GPT-Red adversarial training to strengthen future models against this attack class.

Track how attackers are adapting to this technology.
3 events from the most recent confirmed update back to the earliest known activity.
OpenAI incorporated self-reproduction into GPT-Red attacker objectives so future adversarial training can identify and improve resistance to self-replicating prompt injections.
OpenAI said the self-replicating prompt-injection experiments were limited to simulated tool calls in training and evaluation environments, and that it observed no external impact or real-world incident.
OpenAI reported research showing that prompt injections can both induce unauthorized agent actions and propagate their payload through public output channels, analogous to computer worms. The company demonstrated email- and filesystem-based replication using internal GPT-5.4-mini research checkpoints and a multi-hop Slack scenario involving GPT-5.5.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
2 references tracked. Mallory keeps watching after this page renders.
Map indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.