DeepMind 的 Tom Zahavy 说了个大实话:LLM 搞不了真正的新发现,得靠世界模型。不是泛泛空谈,有论文为证。
Google DeepMind 研究员 Tom Zahavy 在题为“LLMs can't jump”的立场论文中认为,语言模型缺乏创造全新事物的认知机制,因此无法引发科学革命。文章指出,语言模型只能在其训练数据分布内插值,而真正的科学突破需要跳出现有知识框架。作为替代方案,Zahavy 提出“世界模型”可能具备这种能力,因为它们能模拟因果结构和进行反事实推理。该论文并未给出具体基准或数字,而是从哲学和认知科学角度提出理论观点。
Language models can't spark scientific revolutions, but world models might
Can language models spark a scientific revolution? In a position paper titled "LLMs can't jump," Google Deepmind's Tom Zahavy argues they can't. They're missing the cognitive mechanism needed to create something truly new. The article Language models can't spark scientific revolutions, but world models might appeared first on The Decoder .