OpenAI Pauses Frontier RL Training for Two Weeks
OpenAI cited rising risks from capable models and the need for stronger internal safeguards.
TLDR
OpenAI posted that it temporarily paused reinforcement learning training on its latest models intended for deployment. The company said more capable models bring greater risks during internal development and testing, so it hardened and red-teamed research environments while expanding monitoring coverage. Sam Altman stated the pause covers some frontier RL runs to meet alignment, security, and monitoring standards and that it affects further-out releases. He added the company still expects to ship new models soon. Greg Brockman noted the slowdown includes the largest planned frontier RL effort.
Combined views
8.3M
77 Sources, first seen ago