Chollet 明确表示,即使系统在 ARC 3 上饱和,也不代表实现了 AGI,目前对该系统所知仅为 benchmark 分数。ARC 3 测试的是 AGI 系统应有的定性能力——不确定性下的探索、无指令适应、有限数据下的因果世界建模等,但规模极小:游戏时间尺度比真实任务短数个数量级,数据量、建模复杂度和在线学习量也少数个数量级。
Many of you will ask, "if it saturates ARC 3, is it AGI?"
We're not making this claim. All we know about the system so far are its benchmark scores.
When we launched ARC 3, and in every presentation we made about it, we were very insistent on one thing: solving it is not proof of AGI. It's not intended as a finish line.
ARC 3 is testing the right qualitative properties you'd expect of an AGI system -- exploration under uncertainty, adaptation without instructions, causal world modeling from limited data, etc. -- but in small quantities. ARC 3 games are orders of magnitude shorter timescales than real world tasks, and represent orders of magnitude less data, less modeling complexity, less on-the-fly learning.
(Slide below is from a March 2026 presentation)
来源:@fchollet · x.com