Alibaba Qwen团队发布了Qwen3.8-Flash,性能卓越且价格亲民,是DeepSeek V4 Flash和Claude Opus 4.6 Max的性价比之选。
Qwen3.8-Flash,Alibaba Qwen团队的多模态MoE模型,性能优于DeepSeek V4 Flash和Claude Opus 4.6 Max,价格更低。采用Qwen4架构预览,参数125B,N-gram嵌入51B,激活6B。训练和推理成本大幅降低,DeepSWE 1.1评分58.7,SWE-bench Pro 62.5,CoWorkBench 73.9,AndroidWorld 84.5,MathVision 95.7。即将通过QwenCloud API提供,价格0.16/1M输入令牌和0.47/1M输出令牌。
That's absolutely massive You remember how big the DeepSeek V4 Flash release was? Because of the pr...
That's absolutely massive You remember how big the DeepSeek V4 Flash release was? Because of the price/performance ratio? Qwen3.8-Flash is just way better, even cheaper... and open weight!! BETTER than DeepSeek V4 Flash BETTER than Claude Opus 4.6 Max But way cheaper because of the new architecture they're using. This release is incredible. Open weight models are winning! Qwen @Alibaba_Qwen ⚡Meet Qwen3.8-Flash, a multimodal MoE and an early preview of the Qwen4 architecture, now open-weight! The production version Qwen3.8-Flash will be available soon via QwenCloud API at just $ 0.16/1M input tokens and $ 0.47/1M output tokens. 125B parameters + 51B N-gram embeddings, with just 6B activated per token. Unmatched cost-efficiency. What's new: 🥳 - Next architecture: GDN + QSA hybrid attention, Gated Residual, N-gram Embedding & Muon optimizer, serving as a precursor to the architecture used in Qwen4. - Dramatically lower training and inference costs: trained at just 1/9 the cost of Qwen3.7-Plus, while outperforming it across the board with especially strong gains in coding and office tasks. - Strong performance: scoring 58.7 on DeepSWE 1.1, 62.5 on SWE-bench Pro, 73.9 on CoWorkBench, 84.5 on AndroidWorld, and 95.7 on MathVision (with CI). - 262K native context, extensible to 1M with YaRN. We’re also releasing the weights for Qwen3.8-Flash-Next, giving the community an early look at the new architecture we’re exploring for Qwen4.🚀 We can't wait to see what you build with Qwen3.8-Flash!👀👇 - Bl qwen.ai/blog?id=qwen3.… FLgJ - Technical Repo github.com/QwenLM/Qwen3.8… IkQO - Hugging Fa huggingface.co/Qwen/Qwen3.8-F… AABt - ModelSco modelscope.cn/models/Qwen/Qw… NuFG 🔗 View Quoted Tweet 💬 13 🔄 5 ❤️ 139 👀 12733 📊 23 ⚡