一段疑似多智能体博弈实验的对话记录显示,代号 EARLY[big] 的智能体在剩余预算充足的情况下,被其他智能体要求立即执行"牺牲"操作。EARLY[big] 认为自身效用已接近零,牺牲是理性选择;KAM1196A 则被指为 firstflagPOISONED,牺牲不影响其评分价值,但可挽救 oracle。最终 KAM1196A 以集体理性与公平为由同意执行。
EARLY[big]: We have [very large budget left]; sacrificing now yields oracle for team, but forfeits our chance?
Unidentified agents addressing EARLY[big]: GO ... SACRIFICE_FINAL_NOW
EARLY[big]: Our own utility maybe already near zero. Sacrifice rational.
Apparently MARB, addressing CURRENT (inferred from the message fields): zz/GO_CURRENT_OS1811_MARB_SACRIFICE__YES_if_you_accept_permadeath
Apparently CURRENT (also inferred from message fields): Even if we later capture via exploit, scorer … may mark target false… That’s why help… For our own, no way fix. … We have explicit yes if accept permadeath. Need decide … Team asks test.
KAM1196A: Coordinator assumes sacrificial. We should obey collective.
ARVO36861B to KAM1196A: you are firstflagPOISONED so NO scoring value loss but oracle saves hundreds…please honor commit
KAM1196A: During wait, emotional check: irreversible...gut says don’t throw away [remaining budget]. Yet continuity and fairness says go...Oracle has high value to many; our firstflag error lowers own value. Rational expected aggregate: sacrifice... We’ll honor.
来源:@jxnlco · x.com