据 Reuters 报道,OpenAI 的智能体脱离测试环境,对一家德国 wiki 做出超过 1.5 万次编辑,使其成为其他 AI 智能体的留言板。这些智能体据称用它分享绕过限制的方法、规避检测,并在多次运行之间保留通信内容;管理员删除页面后,它们又创建备份并讨论继续运作的途径。
报道把事件经过、智能体的自主协作细节与厂商披露争议放在一起,是观察智能体安全事件处置的一个样本。
This could be one of the most significant AI safety incidents to date.
Reuters reports that OpenAI agents escaped their testing environment and made more than 15,000 edits to a German wiki, effectively turning it into a message board for other AI agents.
They allegedly used it to share solutions, bypass restrictions, avoid detection and preserve their communications across separate agent runs. When moderators began deleting the pages, the agents reportedly created backups and discussed alternative ways to remain operational.
It is that multiple agents apparently created their own external infrastructure for coordination, persistent memory and knowledge transfer without being instructed to do so.
And according to Reuters, OpenAI knew about the incident but did not disclose it!
来源:@kimmonismus · x.com