OpenAI temporarily slowed frontier scaling: two-week pause in RL training and its largest planned frontier RL run still on hold after the Hugging Face incident and Astra's possible Critical cyber threshold
In a company publication dated 18 Aug 2026, OpenAI disclosed it "temporarily slowed the pace of scaling", including "a two-week pause in reinforcement learning (RL) training on our latest models intended for deployment", and stated "Our largest planned frontier RL run remains on hold". Two triggers are named: the OpenAI-Hugging Face incident, and preliminary evidence that its upcoming model Astra "may meet the Critical cybersecurity capability threshold under our Preparedness Framework" (determined 7 Aug). OpenAI says a significant number of Astra workloads remain paused, that the new standards "incurred great cost and delays to frontier research", and that it will evolve the Preparedness Framework.