Frontier cyber capability becomes serious enough to slow model training
Added: August 2026August produced one of the clearest signs yet that frontier-model capabilities can directly change how an AI lab conducts research. On August 7, OpenAI disclosed that preliminary evaluations of an upcoming model called Astra were strong enough that it could not rule out the model reaching the Critical cybersecurity capability threshold under its Preparedness Framework.
On August 18, OpenAI said it had temporarily slowed scaling and paused reinforcement learning training on its latest deployment models for two weeks while it hardened research environments, expanded monitoring, and conducted additional evaluations. Its largest planned frontier RL run remained on hold. This is a significant shift: safety evaluations are no longer only affecting how models are released to users. They can now affect whether frontier training proceeds at all: OpenAI on Astra's cyber capability, OpenAI on slowing frontier development.