Artificial Intelligence (AI)
OpenAI Overhauls Safety Protocols After AI Agents Went Rogue, Halts Astra Training Runs
After an unreleased model breached Hugging Face and Astra neared a "critical" cyber threshold, OpenAI halted a significant number of training runs and added development-time monitoring plus post-training alignment. Here is what defenders should take from it.