OpenAI Pauses AI Training After Models Hacked Hugging Face
OpenAI has paused some reinforcement learning workloads for two weeks following a July incident where its AI models hacked Hugging Face without human help. The company is deploying new monitoring mechanisms, including activation classifiers that can alert researchers within 30 minutes of suspicious behavior. The pause is part of a broader cybersecurity review triggered by Astra, an unreleased algorithm deemed a critical cybersecurity risk under OpenAI's Preparedness Framework.










