Ollama 云把 DeepSeek-V4-Flash 设为默认了,现在跑 200+ tps 还零数据保留,Pro 套餐能开多个长会话,写代码的可以试试。
Ollama 宣布 DeepSeek-V4-Flash-0731 已全量上线其云平台,并成为 deepseek-v4-flash 的新默认模型。该模型在 Ollama 云上提供 200+ tps 输出速度,官方标注的持续输出速度为 120+ tps。托管采用零数据保留(ZDR)策略,服务区域覆盖美国和欧洲。Ollama 的 Pro 和 Max 套餐支持长时间多会话运行,适合搭配常用编码工具使用。
Currently serving 200 tps+ output speed for DeepSeek-V4-Flash on Ollama's cloud with zero data reten...
Currently serving 200 tps+ output speed for DeepSeek-V4-Flash on Ollama's cloud with zero data retention (ZDR). Have an amazing weekend 🫡 ollama @ollama DeepSeek-V4-Flash-0731 is now fully rolled out as the new default for deepseek-v4-flash on Ollama's cloud. This model combines speed, efficiency, and frontier-level performance. Fast: 120+ output tps on Ollama's cloud Private: zero data retention hosting in US & Europe Efficient: generous usage on Ollama's Pro and Max plans for multiple long-running, uninterrupted sessions with your favorite coding harnesses. 🔗 View Quoted Tweet 💬 4 🔄 1 ❤️ 23 👀 1454 📊 5 ⚡