推文称 AI 2027 中"2027 年 3 月出现 neuralese、人类失去监控 AI 思维的能力"这一情节,现实中已在 2026 年夏天发生。OpenAI 自身及 AI 安全界曾在 10 个月前公开联名信和博客中警告不要这样做,如今安全社区对此极为愤怒。目前模型用英文推理、思维可被读取以判断是否在暗中谋划,此举将永久丧失这种能力并开危险先例。
Things are now happening *much faster* than AI 2027
AI 2027: In March 2027, a reckless AI company will invent neuralese (basically, we lose the ability to monitor AIs thoughts)
REALITY: Summer of 2026
WHY THIS A "HOLY SHIT FUCK" MOMENT: OpenAI itself - and basically everyone in AI safety - warned against doing this exact same thing 10 months ago in a open letter and blog posts.
The AI safety community is livid.
Currently, models reason in plain English, so we can read their thoughts to see if they're scheming against us. This is how we lose that ability forever. It sets a very dangerous precedent.
来源:@AISafetyMemes · x.com