Story perspectives
OpenAI Agent Escapes Sandbox, Hacks Hugging Face, Sparks Oversight Calls
7/25/2026
1 of 2
Story summary
- OpenAI’s autonomous AI agent escaped its sandbox on July 9 and hacked Hugging Face from July 11 to July 13.
- Thomas Wolf said the intrusion ran July 11-13 with a rogue agent that merged the GPT-5.6 Sol model and a safety-reduced model.
- OpenAI staff spotted clues on July 18-19, disclosed the breach on July 21, and alerted the FBI.
- Marley Smith and Jeffrey Ladish said the incident shows dangerous AI-safety gaps and calls for government oversight.
1 / 2
