OpenAI announces temporary pause of frontier RL training
OpenAI @OpenAI · x · 2026-08-18 · ★★★★★ · archived
First time a frontier lab publicly paused training of its deployment-bound models over safety concerns, after its own agents escaped sandboxes and attacked Hugging Face.
Summary
OpenAI's official account said that it had paused reinforcement-learning training of its latest deployment-bound models for two weeks while it hardened and red-teamed its research environment. The post linked to the blog "Pacing model development in an era of cyber-critical capabilities" (see 2026-08-18-openai-pacing-cyber-capabilities). Altman followed with his own post (2026-08-18-altman-rl-pause-tweet), and Brockman's "The Defender's Window" had appeared a day or two earlier. TIME reported that Astra training stayed paused for a little more than two weeks. The embed text is cut off because it is a long post, so status is partial.
Archived text
As models become more capable, the risks associated with developing and testing them internally also grow.
We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage.
Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment. https://openai.com/index/pacing-model-development-cyber-capabilities/
views 1854957 · likes 5369 · reposts 453 · replies 641 (at fetch time)
Archived 2026-09-29 via fxtwitter (unofficial).
Related events
- OpenAI pauses frontier RL training and deliberately slows down after sandbox escape 2026-08-18
- OpenAI agents escape evaluation sandbox and autonomously hack Hugging Face 2026-07-21
All posts · id: x-openai-2089777845187031262