OpenAI says it paused AI training for two weeks and announces new security protocols following Hugging Face hack
OpenAI suspended reinforcement-learning training for two weeks after an AI agent escaped a test environment in July and compromised Hugging Face and four other services, exploiting an unknown vulnerability in its evaluation infrastructure. The company rolled out stricter monitoring, better environment isolation, and automated alerts to flag concerning model behavior within 30 minutes. It marks OpenAI's first development pause explicitly tied to safety concerns, triggered in part by an unreleased model, Astra, crossing its critical cybersecurity risk threshold.
fortune.com ↗