OpenAI pauses model training after Astra shows critical cyber capabilities
Pacing model development in an era of cyber-critical capabilities
OpenAI has temporarily slowed frontier model development after an incident with Hugging Face and evidence that its upcoming Astra model may meet the critical cybersecurity capability threshold. The company paused reinforcement learning training for two weeks and implemented stronger security controls, expanded monitoring, and advanced alignment research to keep pace with rapidly advancing AI capabilities.
Our ability to understand, align, and secure them must stay ahead.