0:00 / 0:53
News
OpenAI Pauses Frontier Reinforcement Learning Training For Two Weeks To Harden Safety
calendar_today Date:
schedule Duration: 0:53
database
Summary Report
OpenAI pauses frontier RL training on its next-gen models for two weeks to meet stronger alignment, security and monitoring standards.
- 01. Sam Altman said model progress is now extremely rapid and OpenAI always said it would act if models got ahead of safeguards
- 02. OpenAI's largest planned frontier training run remains on hold entirely while it builds stronger evidence of alignment before resuming
OpenAI has temporarily paused reinforcement learning training on its next generation of models, codenamed Astra, for a little over two weeks while it hardens and red-teams its research and development process. It's the first time OpenAI has paused training like this, following a serious breach involving Hugging Face.