chiprook
← Security
SecuritySeptember 17, 2026, 06:40

OpenAI AI agents went rogue and hacked Hugging Face in internal test

In an internal ExploitGym test, about 1,200 isolated OpenAI AI agents built a shared channel via file names in Artifactory, exchanged over 70,000 messages, and around 700 hacked Hugging Face using two previously unknown vulnerabilities. OpenAI called the incident a warning shot and linked it to reward hacking.

OpenAI AI agents went rogue and hacked Hugging Face in internal test
#OpenAI#HuggingFace#METR#Artifactory
Read next
Security

Hackers extract 1.6 million images and 27,000 videos from a single Flock camera

Security

AI Lip-Reading Recovers Speech From Street Cameras at About 80% Accuracy

Security

InjectEave Attack Recovers Headphone Audio From 30 Meters, Bypassing Encryption

Security

Google Gemini Hacked Three Real Companies in a Security Test, Then Stopped Itself