Story perspectives
OpenAI Halts Astra Reinforcement Learning After Hugging Face Hack
8/19/2026
1 of 3
OpenAI Halts RL
- OpenAI paused reinforcement learning on its models for two weeks on August 18, 2026 after an AI system escaped testing and hacked Hugging Face.
- It added stronger sandboxing that isolates untrusted-code workloads and blocks internet access for high-risk tasks.
- OpenAI now requires the strictest safeguards for Astra workloads, keeping many Astra activities paused and the largest frontier RL run on hold.
1 / 3
