@EMostaque· @EMostaque · X·· 2026-08-19精选AI 评分74
AI 导读
OpenAI 表示已暂停部分前沿 RL 训练,以确保满足当前新能力水平所要求的对齐、安全与监控标准。Sam Altman 称模型进展极快,若能力超出安全与对齐的节奏就会采取行动,并认为安全信心将越来越决定 AI 进展的节奏,同时希望整个领域协调共享安全标准,在此之前会单方面行动。Emad Mostaque 公开赞赏这一做法,认为眼下正发生奇怪甚至危险的事情,而现有系统尚未做好准备,并提到非前沿模型已足以改变生活。
推荐理由
OpenAI 因安全与对齐标准暂停部分前沿 RL 训练,这条公开表态可作为观察安全节奏如何影响模型进展的样本。
正文
I would like to publicly commend @sama and @OpenAI for this taking it at face value.
It is very clear that strange and perhaps dangerous things are happening and our systems are not ready for this.
Models below frontier are competent enough to change lives so lets optimse https://t.co/tnQmEFL0oN
We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment. We care very deeply about AI safety. We believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime. We expect confidence in safety to increasingly set the pace of AI progress. We are optimistic about the alignment work we are doing, and we remain committed to making frontier capabilities widely available. https://t.co/51kvKfbfrO在 X 查看被引用的帖子
来源:@EMostaque · x.com