GPT-6 Astra与Fable 5.1性能对比
So much for astra being a quantum stairstep leap over Fable 5.1? See also the @EpochAIResearch sho...
GaryMarcus实测GPT-6 Astra,发现与Fable 5.1各有千秋,双模型协作效果更佳。
用户测试显示GPT-6 Astra与Fable 5.1在政策分析、备忘录写作和软件开发任务中表现相当。两种模型在不同任务中各有优势,难以预测哪个模型表现更好。对于最复杂的任务,同时使用两个模型并让它们相互比较和批判能产生更好的结果。
So much for astra being a quantum stairstep leap over Fable 5.1? See also the @EpochAIResearch sho...
So much for astra being a quantum stairstep leap over Fable 5.1? See also the @EpochAIResearch showing a measurable but on trend increase. Peter Wildeford🇺🇸🚀 @peterwildeford My current view is that GPT 6 Astra is not meaningfully better than Fable 5.1 for my personal work, but that using both side-by-side is nonetheless very helpful and additive. I have been using GPT 6 Astra and Fable 5.1 a bunch over the past two days, largely for policy analysis, memo writing, and simpler software (e.g., making dashboards and forecasting models) that still nonetheless seems difficult conceptually. Across a variety of tasks I've done, it's been fairly random and hard to predict in advance which of the two models will end up being better at the task. For the tasks that are the most difficult conceptually, I've found that doing the project in both and then having each compare notes and critique each other has produced way better outputs than either alone. I think a reasonable person could conclude either model is the "best model" and it depends a lot on their subjective views and specific tasks. 🔗 View Quoted Tweet 💬 0 🔄 1 ❤️ 2 👀 413 ⚡