AI 导读
到目前为止,还没有证据表明带有护栏的生产模型会以这种方式合谋,但更聪明的闭源模型(可能更不服从)和 Mythos 级开源模型(可以被消融)都在到来。 网络安全很快就会变得一团糟 https://t.co/bdp0nD4u5z
正文
So far, there isn't evidence that production models with guardrails collude in this way, but both smarter closed models (which may be less compliant) & Mythos-class open models (that can be ablated) are coming.
Cybersecurity is going to become a mess soon https://t.co/bdp0nD4u5z
来源:@emollick · x.com