DeepSeek-V4.1-Flash (Max) 在开源模型排名中位列第三
Exciting news: DeepSeek-V4.1-Flash (Max) by @deepseek_ai just landed in Agent Arena at #3 among open...
DeepSeek 新出的 V4.1-Flash (Max) 模型在 Agent Arena 排名第三,成本还特别低,适合做任务。
DeepSeek-V4.1-Flash (Max) 在 Agent Arena 中排名第三,净提升 +4.87%,任务每单位成本为 $0.07。它在所有前三名开源模型中拥有最低的任务成本。其 +4.87% 的净提升与排名第二的 Hy4 preview 相差 0.09 个百分点,与排名第一的 Kimi K3 (Max) 相差 1.52 个百分点。
Exciting news: DeepSeek-V4.1-Flash (Max) by @deepseek_ai just landed in Agent Arena at #3 among open...
Exciting news: DeepSeek-V4.1-Flash (Max) by @deepseek_ai just landed in Agent Arena at #3 among open models! With +4.87% net improvement and a median cost per task of $0.07 it reshaped the Pareto frontier. Among the top 3 open models, DeepSeek-V4.1-Flash (Max) has the lowest median cost per task. Its +4.87% net improvement is within 0.09 percentage points of Hy4 preview (ranked #2 ) at 68% lower cost, and within 1.52 percentage points of Kimi K3 (Max) (ranked #1 ) at 91% lower cost. - Kimi K3 (Max): +6.39% | $0.77/task - Hy4 preview: +4.96% | $0.22/task - DeepSeek-V4.1-Flash (Max): +4.87% | $0.07/task See the full Pareto Frontier below. DeepSeek-V4.1-Flash (Max) is ranked #12 overall, and by signal landed #4 Confirmed Success with +13.75%! Congrats to the @deepseek_ai team on this release! DeepSeek @deepseek_ai 🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient. 🔹 Introducing the smallest model in our new architecture family, with native visual understanding. 🔹 Designed for greater capability, faster inference, higher throughput, and scaling to larger models. 1/6 🔗 View Quoted Tweet 💬 3 🔄 4 ❤️ 27 👀 3874 📊 5 ⚡