Cognition 发布 SWE-2 模型,在代码生成任务上表现接近 Fable 5.1
Just amazed by how specialized models keep pushing the Pareto frontier. Just watch that massive co...
Cognition 推出了 SWE-2,这个模型在代码生成上很厉害,和 Fable 5.1 差不多,但便宜很多。
Cognition 公司的 SWE-2 模型在 FrontierCode 1.1 Main1 基准测试中取得了 50.0% 的成绩,仅比 Fable 5.1 低 1 分,同时成本降低了 64%。该模型通过扩展强化学习至数千亿参数,在能力和成本之间取得了新的平衡。
Just amazed by how specialized models keep pushing the Pareto frontier. Just watch that massive co...
Just amazed by how specialized models keep pushing the Pareto frontier. Just watch that massive cost reduction. SWE-2 achieves 50.0% on FrontierCode 1.1 Main1, within one point of Fable 5.1, while being 64% cheaper. Cognition @cognition Introducing SWE-2, our closest model yet to the frontier. On leading evals, it scores on par with recent frontier models – at up to 70% lower cost. We scaled RL to multiple trillions of parameters, with a refined recipe that pushes the Pareto curve on both capabilities & cost. 🔗 View Quoted Tweet 💬 2 🔄 3 ❤️ 23 👀 2149 📊 5 ⚡