Mercury 2.5 Preview来了,推理速度超1100 tokens/sec,还支持并行工具调用,比普通模型快多了。
Mercury 2.5 Preview推理模型在OpenRouter平台发布,速度达1,107 tokens/sec。该模型支持并行token生成和并行工具调用功能。模型具有可调节推理能力和schema-aligned JSON输出。专为低延迟工作负载设计。
1/ The fastest reasoning LLM is now live exclusively on OpenRouter. Mercury 2.5 Preview from @_ince...
1/ The fastest reasoning LLM is now live exclusively on OpenRouter. Mercury 2.5 Preview from @_inception_ai reaches 1,107 tokens/sec through parallel token generation, with tunable reasoning, parallel tool calls, and schema-aligned JSON. Built for latency-sensitive workloads. Your browser does not support the video tag. 🔗 View on Twitter 💬 8 🔄 6 ❤️ 36 👀 2914 📊 13 ⚡