ARC Prize测试显示OpenAI模型依赖外部框架
so it was the harness not the model. more sneaky games with unacknowledged domain-specific neurosym...
GaryMarcus揭露OpenAI模型测试差异,显示实际效果依赖外部框架而非模型本身。
ARC Prize基准测试中,OpenAI模型在相同测试下得分差异显著。模型本身得分为99.9%,但基准测试软件返回结果为62.7%。ARC Prize组织表示未声称实现AGI,指出不同框架导致得分差异。
so it was the harness not the model. more sneaky games with unacknowledged domain-specific neurosym...
so it was the harness not the model. more sneaky games with unacknowledged domain-specific neurosymbolic AI doing the real work. TNW @thenextweb OpenAI published 99.9%. The benchmark's own software returned 62.7%. Same model, same test, different scaffolding around it. ARC Prize printed both numbers and says it is not claiming AGI. thenextweb.com/news/openai-as… 🔗 View Quoted Tweet 💬 10 🔄 6 ❤️ 57 👀 5244 📊 14 ⚡