Pacing model development in an era of cyber-critical capabilities
Advanced cybersecurity capabilities in AI increase the risk of unauthorized access or destructive attacks if systems are not properly secured. These measures protect internal and external networks from potential model-driven security breaches.
- OpenAI paused RL training for two weeks on its latest models intended for deployment while hardening and red-teaming research environments.
- Preliminary evidence indicates Astra may meet the Critical cybersecurity capability threshold under OpenAI's Preparedness Framework, determined on August 7, 2026.
- OpenAI's largest planned frontier RL run remains on hold while smaller-scale training and evaluations continue.
- OpenAI expanded chain-of-thought monitoring and now aims to issue alerts within 30 minutes of concerning model activity, with monitoring overhead estimated at roughly 20% of monitored inference compute.