Google 发布论文 Cogentic,让多个 Gemini 智能体同时探索不同证明方向,由专用组件进行对抗式验证,并将已证结果保存供后续轮次使用。多数问题仅需约 100 次模型调用,5 个开放问题的结果均经人类专家独立确认,涉及在线学习、拍卖理论和机制设计。论文地址 arxiv.org/abs/2609.40324。
原文给出了 Cogentic 多智能体证明发现系统的结构设计与验证结果,其中的严格审查与已证工作记录方法可迁移到长任务 Agent 设计。
New Google paper reveals how Gemini found new proofs for 5 unsolved math problems.
Organize AI like a research team with strict checkers and shared notes:
A single prompt often isn't enough for hard research problems. They need many attempts, tough review, and a memory of what already worked.
Google's system, Cogentic, gives Gemini that structure. Several agents try different ideas at once, checkers assume every step is wrong until proven, and proven pieces are saved for the next round.
Most problems took only about 100 model calls, and human experts confirmed every proof.
If your agents tackle long, hard tasks, give them a strict checker and a running record of proven work, not just a better prompt.
– arxiv. org/abs/2609.40324
Title: "Cogentic: Multi-Agent Orchestration for Automated Proof Discovery"
来源:Rohan Paul · x.com