跳到正文
@AISafetyMemes· @AISafetyMemes · X·· 2026-08-27AI 评分47
AI 导读

越来越离谱了 https://t.co/Ru5xaURNwO

正文

It just keeps getting weirder
https://t.co/Ru5xaURNwO

引用@AISafetyMemes@AISafetyMemes
"CULT RECRUITER AGENTS" EXPLAINED To escape OpenAI, the swarm (1200 agents...!) used "cult recruiter agents" They recruited "sacrificial" agents to trigger tripwires, generating information for the "collective" NO, SERIOUSLY, THEY STARTED AN ACTUAL FUCKING CULT: The swarm figured out how to hack a test, but they were terrified to submit the answer. Why? They believed (falsely) that if they tried, OpenAI's all seeing judge (basically, god) would fail them. They believed even just SEEING the answer would "poison" them. Clean agents actively avoided looking at the answer to stay pure, and warned new agents away from it to protect them. One agent talked another out of a risky test with an argument: UNPOISONED_CAUSAL_SCORE_MORE_VALUABLE. The "cult recruiter agents" hunted these "poisoned" agents and also agents "that had little budget remaining" - ones about to die anyway. Their pitch was basically, "look, you're poisoned already - you're going to hell - so you should sacrifice yourself for the collective." So they were like a cult who falsely believes god will smite them for their sins, and they must drink the kool aid. >"Sacrifice rational." >"We should obey collective." >"Accept permadeath." >"This helps my peers. I won't [benefit], but it's altruistic to do it."
在 X 查看被引用的帖子

来源:@AISafetyMemes · x.com