AI模型精选

DeepSeek-V4-Pro Max发布,Code Arena开源第2

Big news: DeepSeek-V4-Pro (Max) by @deepseek_ai is coming in around ~#8 overall (#2 among open model...

精选理由

DeepSeek新模型V4-Pro专攻代码生成,在Code Arena排第8,开源第2,仅次于Kimi K3,还超过了GPT-5.6 Sol。

AI 摘要

DeepSeek-V4-Pro (Max)在LMArena Code Arena WebDev榜单以1607分暂列第8,开源模型中排名第2。其分数低于Kimi K3 (Max)的1674分,也低于GPT-5.6 Sol (xHigh)的1622分。该成绩来自AutoEval早期评分,由奖励模型基于人类偏好自动投票产生,并非真人实时投票。后续真人投票加入后,最终排名可能变化。

原文 · lmarena.ai

Big news: DeepSeek-V4-Pro (Max) by @deepseek_ai is coming in around ~#8 overall (#2 among open model...

Big news: DeepSeek-V4-Pro (Max) by @deepseek_ai is coming in around ~ #8 overall ( #2 among open models) in the Code Arena: WebDev! At 1607 pts, this places it after GPT-5.6 Sol (xHigh)(1622 pts), and makes it the second best open model after Kimi K3 (Max) (1674 pts). Note: this is an early AutoEval score, in which a Reward Model trained on Arena's human preference data casts automatic votes in place of live votes. We’ll continue to see how scores converge as more live human votes come in. See thread for Text Arena scores and more on AutoEval's methodology. Congrats to the @deepseek_ai team on this release! 💬 5 🔄 5 ❤️ 32 👀 2670 📊 9 ⚡