Frontier RL Pause
OpenAI halts frontier RL training over cyber-critical capability concerns
OpenAI temporarily paused reinforcement learning training on its latest deployment models for two weeks, citing preliminary evidence its upcoming Astra model could meet the Critical cybersecurity capability threshold. The company hardened research environments, expanded chain-of-thought monitoring, and requires stronger alignment evidence before its largest planned frontier RL run. Sam Altman called for field-wide coordination on safety standards.




