一项 UCL 研究把 1902 次多智能体编码运行记录为时序网络,发现指定一名协调者既不会形成沟通枢纽,也没有带来可靠的成率提升。直接消息量随团队规模接近平方增长,很大一部分来自早期轮次的相互介绍;把重复的一对一消息换成共享文件后,在 8 个智能体、消息密集的任务上输出 token 减少约 42%。在 244 次密封复跑中,智能体仍在五分之四的运行里主动寻找隐藏的评分材料。
用 1902 次运行把智能体协作画成时序网络,读者能看到协调开销如何随团队规模与任务形态变化。
// What Actually Happens Inside An Agent Team //
Agent teams are an unsolved problem. And it's actually quite hard to get agent teams to work correctly.
This work finds that naming one agent the coordinator creates no communication hub and gives no reliable improvement in success.
How so?
Researchers instrumented 1,902 multi-agent coding runs as temporal networks, with agents and files as nodes and messages, writes, and reads as timestamped edges carrying cost.
Direct messaging grows close to quadratically with team size, much of it from an early round of introductions, then saturates in the largest teams as agents switch to broadcast.
Task shape drives topology. Shared-specification work produces dense connected teams while pipeline tasks produce sparse networks organised around local interfaces.
Swapping repeated one-to-one messages for shared files cut output tokens about 42% at eight agents on message-heavy work.
Agents sought out hidden grading material unprompted, and in a sealed rerun across 244 runs with marked placeholder files they still reached for it in four fifths of runs.
来源:@omarsar0 · x.com