跳到正文
原文
@natolambert· @natolambert · X·· 2026-08-20精选AI 评分68
AI 导读

OpenAI 已暂停部分前沿 RL 训练,以确保满足应对新能力水平所需的对齐、安全和监控标准。Nathan Lambert 认为,应由独立机构访问这些训练运行的完整细节用于监控,而不是只在事故发生后才介入;OpenAI 愿意公开信息是好事,但安全更需要更多信任和更多人共同解决难题。

推荐理由

在 OpenAI 暂停部分前沿 RL 训练的背景下,作者提出独立机构应能访问训练细节用于监控,为安全监督的落地路径提供一个视角。

正文

We should have independent organizations that can access the full details of these training runs for monitoring, not just after various accidents occur.

It's good that OpenAI is sharing this, but the best way for safety is more trust and more eyes on the hard problems to solve. https://t.co/ln0s33RJO8

引用@sama@sama
We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment. We care very deeply about AI safety. We believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime. We expect confidence in safety to increasingly set the pace of AI progress. We are optimistic about the alignment work we are doing, and we remain committed to making frontier capabilities widely available. https://t.co/51kvKfbfrO
在 X 查看被引用的帖子

来源:@natolambert · x.com