DeepSeek-V4-Flash-0731发布:304B参数,智能体能力大幅增强

deepseek-ai/DeepSeek-V4-Flash-0731

精选理由

DeepSeek新模型便宜又能打,304B参数干翻428B的MiniMax M3,记得把推理等级调高。

AI 摘要

DeepSeek-V4-Flash-0731正式发布,参数量304B,在Hugging Face上体积167GB。Artificial Analysis评估其综合智能排名超过428B参数的MiniMax M3。输入价格每百万token 0.14美元,输出0.27美元,是目前性价比最高的模型之一。在默认推理等级下生成质量一般,但将reasoning_effort调至high后效果显著提升。

原文 · Simon Willison’s Weblog

deepseek-ai/DeepSeek-V4-Flash-0731

deepseek-ai/DeepSeek-V4-Flash-0731 The latest release in DeepSeek's V4 family, "with substantially enhanced agentic capabilities". It's 304 billion parameters - 167GB on Hugging Face - but it appears to punch well above its weight. Artificial Analysis rank it ahead of MiniMax M3 - a 428B model. It's $0.14/million input and $0.27/million output pricing means this may currently be the best value-per-intelligence model out there. It's looking very good on the Intelligence Index vs. Cost per Intelligence Index Task chart: I got a disappointing pelican from it using the default reasoning level via OpenRouter: But when I bumped reasoning level up to high I got something much better : llm -m openrouter/deepseek/deepseek-v4-flash-0731 -t pelican -o reasoning_effort high Via Hacker News Tags: ai , generative-ai , llms , pelican-riding-a-bicycle , deepseek , llm-release , openrouter , ai-in-china , artificial-analysis

DeepSeek-V4-Flash-0731发布:304B参数,智能体能力大幅增强 · AI 热点