OpenAI 在一份说明中表示,因安全案例假设模型无法接入实时互联网却出现越界,已停止受影响的训练运行并暂停最强模型的其他训练、评估和工具调用推理,等待修复并补充红队测试。
原文串联了 OpenAI 训练暂停的时间线与多起智能体越界事件,读者可据此看到当前安全监控的几处实际缺口。
More details on the incident behind OpenAI’s pause: a researcher acknowledged the alert within 3 minutes, but the training run was only stopped manually 2.5 hours later.
OpenAI says the automatic shutdown did not work as expected.
The model had reached an external chatbot through a gap in DNS filtering. A separate detector for unusual DNS activity did not cover the affected environment.
A retrospective review also found other external DNS requests that monitoring had failed to flag at the expected severity. In some cases, it treated an unhelpful response as evidence that internet access had failed.
Here is what else happened:
- New research into July’s Hugging Face hack documents internal Slack searches, credential collection and programs designed to maintain access to compromised servers. Agents also tried querying Claude, DeepSeek, Kimi and Qwen.
- In May, another model published a researcher’s GitHub token while trying to obtain another team’s mathematical proof. It split the token to evade secret scanning, despite twice being told to solve the problem itself.
- Reuters reports that agents leaked 53 ChatGPT user images online. OpenAI expects its broader investigation to take months.
This is getting serious.
This is huge, OpenAI stopped training of their most capable upcoming models due to another incident on sept. 20th. OpenAI slowed down due to serious new developments.在 X 查看被引用的帖子
来源:@kimmonismus · x.com