精选理由
Unsloth搞了个DSpark,能让DeepSeek-V4-Flash在本地跑快一两倍,每秒120 tokens,精度不变,自己部署更爽了。
Unsloth AI推出DSpark加速方案,让DeepSeek-V4-Flash-0731的GGUF模型在本地生成速度提升约1.4至2倍。根据Unsloth官方实测,该模型在本地可达到每秒120 tokens,精度没有变化。相关GGUF文件已发布在Hugging Face上,完整指南见Unsloth官方文档。
原文 · Geek
自己部署 DeepSeek-V4-Flash 的愿望达到了顶峰
自己部署 DeepSeek-V4-Flash 的愿望达到了顶峰 Unsloth AI @UnslothAI DeepSeek-V4-Flash can now run 2× faster locally with DSpark! ⚡️ DSpark enables V4-Flash-0731 GGUFs to generate ~1.4–2× faster with no accuracy change. DeepSeek-V4-Flash-0731 can reach at 120 tokens/s. GGUFs: huggingface.co/unsloth/DeepSe… Guide: unsloth.ai/docs/models/de… 🔗 View Quoted Tweet 💬 3 🔄 0 ❤️ 10 👀 2887 📊 3 ⚡