跳到正文
原文
@kimmonismus· @kimmonismus · X·· 2026-08-19精选AI 评分73
AI 导读

OpenAI 暂停了最新部署模型的强化学习训练两周,最大规模的前沿 RL 运行仍处搁置状态,期间只开展小规模训练与评估,以验证模型行为、安全防护和对齐证据。作者称初步结论显示即将推出的 Astra 模型可能已达到 OpenAI 的 Critical 网络安全阈值,并提及 OpenAI 与 Hugging Face 事件,认为 Astra 短期内不会发布。

推荐理由

OpenAI 暂停最新部署模型的 RL 训练并披露 Astra 或触及 Critical 网络安全阈值,可据此了解前沿模型发布前的安全审查流程。

正文

Not looking good for a soon GPT-Astra-release: OpenAI paused reinforcement learning on its latest deployment models for two weeks, and its largest planned frontier RL run remains on hold.

The company is running smaller-scale training and evaluations while it tests model behavior, safeguards, and evidence of alignment.

The decision follows preliminary findings that its upcoming Astra model may have reached OpenAI’s “Critical” cybersecurity threshold, alongside the OpenAI–Hugging Face incident.

"While some Astra training and evaluations meet those requirements, a significant number of workloads remain paused until they are fully migrated and enhanced to meet the new security bar"

Dont think Astra will be released any time soon. The question is: will china catch up in the meantime?

引用@OpenAI@OpenAI
As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment. https://t.co/ecbMMmVoox
在 X 查看被引用的帖子

来源:@kimmonismus · x.com