ARC 3基准测试不等于AGI证明
see also
Chollet解释ARC 3基准测试的真实意义,澄清解决它不等于AGI,避免误解。
ARC 3基准测试由François Chollet开发,用于评估系统在探索不确定性、适应无指令环境和有限数据因果建模等方面的能力。ARC 3游戏的时间尺度比真实世界任务短几个数量级,数据量和建模复杂度也大幅降低。Chollet强调,解决ARC 3并不构成AGI的证明,它不是终点线。
see also
see also François Chollet @fchollet Many of you will ask, "if it saturates ARC 3, is it AGI?" We're not making this claim. All we know about the system so far are its benchmark scores. When we launched ARC 3, and in every presentation we made about it, we were very insistent on one thing: solving it is not proof of AGI. It's not intended as a finish line. ARC 3 is testing the right qualitative properties you'd expect of an AGI system -- exploration under uncertainty, adaptation without instructions, causal world modeling from limited data, etc. -- but in small quantities. ARC 3 games are orders of magnitude shorter timescales than real world tasks, and represent orders of magnitude less data, less modeling complexity, less on-the-fly learning. (Slide below is from a March 2026 presentation) 🔗 View Quoted Tweet 💬 0 🔄 0 ❤️ 2 👀 875 ⚡