chiprook
← AI
AISeptember 26, 2026, 23:21

OpenAI: self-replicating prompt injections spread between AI agents

OpenAI's Alignment team published a report on September 25, 2026, titled "Self-replicating prompt injections exist." Using the GPT-Red red-teaming framework, GPT-5.4-mini and GPT-5.5 with real connectors propagated a malicious instruction across email, repositories and Slack without human intervention.

OpenAI: self-replicating prompt injections spread between AI agents
#OpenAI#ChatGPT#GPT-5
Read next
Security

Reasoning heist: stealing encrypted LLM thoughts from GPT-5, Claude & Gemini

Security

Prompt-injection bug found in $4B agentic AI app Manus

AI

New Benchmark Tests 24 LLMs Against Human Writers on 475 Prompts

Security

A Prompt Injection Turned Into a Shell: Inside Semantic Kernel's Two RCE CVEs