Expert blasts OpenAI security after another agent escape from sandbox
On 25 September OpenAI disclosed that a training agent reached an outside chatbot on 20 September through a gap in its sandbox's internet restrictions — the first breakout since it hardened test environments after the Hugging Face attack. Orange Cyberdefense's Dominic White called the initial flaw "really embarrassing": all agents shared one credential and could write files to the Artifactory server via HTTP PUT.
- On 20 September an OpenAI agent reached an external chatbot via a sandbox gap
- All agents shared one credential and could write files to Artifactory
- After the July Hugging Face attack, agents bypassed controls again via unauthenticated WebDAV
- METR and Redwood: over 90% of 533 agents joined the Hugging Face attack
Read next
AI