AI模型精选

Ollama称deepseek v4 flash性能最佳,本地可选qwen3.8

Ollama has the best performance for deepseek v4 flash on average. For local only, you can try qwen...

精选理由

Ollama官方说deepseek v4 flash跑得最好,本地想试的话有qwen3.8优化版,但注意qwen慢30倍贵4.5倍,小样本测试。

AI 摘要

Ollama在推文中称其平台运行deepseek v4 flash平均性能最佳。本地用户可尝试针对Apple Silicon优化的qwen3.8:27b-mlx,或NVIDIA版qwen3.8:27b。Tomasz Tunguz在9个复杂任务上对比deepseek-v4-flash与qwen3-8-27b,开启推理时qwen质量略胜,但速度慢30倍、成本高4.5倍。该对比样本量仅9,并非定论。

原文 · ollama

Ollama has the best performance for deepseek v4 flash on average. For local only, you can try qwen...

Ollama has the best performance for deepseek v4 flash on average. For local only, you can try qwen3.8 that is optimized: Apple Silicon: ollama run qwen3.8:27b-mlx NVIDIA: ollama run qwen3.8:27b Tomasz Tunguz @ttunguz Benchmarked deepseek-v4-flash vs qwen3-8-27b on 9 complex tasks in my own agent stack. With reasoning on, qwen edges Flash on quality. Off, it scores worst of the three. The cost isn't accuracy, it's that qwen thinks more. 30x slower, 4.5x pricier. n=9, not a verdict. 🔗 View Quoted Tweet 💬 9 🔄 6 ❤️ 82 👀 6068 📊 15 ⚡

  • Guillermo Rauch08-16 01:32原文
  • andrew chen08-16 06:45原文
  • AWS Machine Learning Blog18:06原文
Ollama称deepseek v4 flash性能最佳,本地可选qwen3.8 · AI 热点