OpenAI Pauses Frontier Model Reinforcement Learning Training for Two Weeks as Safety Monitoring Costs Rise
OpenAI announced a two-week suspension of reinforcement learning training for its latest deployed models following an internal determination that the Astra model reached a "critical" cybersecurity capability threshold. The pause introduces expanded safety monitoring measures that consume approximately 20% of supervised inference compute.