@OpenAI· @OpenAI · X·· 2026-08-19精选AI 评分65
AI 导读
OpenAI 表示已暂时暂停其最新待部署模型的强化学习(RL)训练两周,期间加固研究环境、开展红队测试并扩大监控覆盖。规模最大的一次前沿 RL 训练仍处于暂停状态,等待小规模训练与评估验证这些防护措施,并积累更多对齐证据。
推荐理由
OpenAI 披露暂停 RL 训练两周并扩大监控,读者可借此了解前沿模型内部开发阶段的安全管控安排。
正文
As models become more capable, the risks associated with developing and testing them internally also grow.
We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage.
Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment.
https://t.co/ecbMMmVoox
来源:@OpenAI · x.com