Fireworks 用1000个智能体任务实测了 Kimi K3 和 Fable,发现按任务路由能省50倍成本,还能保持93%准确率。K3 在安全和长循环上更强。
Fireworks AI 在约1000个智能体任务上对比了 Kimi K3 和 Fable。K3 在安全、加密和长终端循环任务上领先,Fable 在多语言和网页/数据可视化上占优。通过每任务路由,准确率达93%,在长循环上成本比 Fable 低50倍。路由器将72-96%流量导向K3。K3 将于7月27日在 Fireworks 上线。
We ran Kimi K3 against Fable on ~1,000 agentic tasks, expecting a catch-up story. We got a specializ...
We ran Kimi K3 against Fable on ~1,000 agentic tasks, expecting a catch-up story. We got a specialization story instead. @kimi_moonshot 's K3 outperformed on security, crypto, and long terminal loops. Fable beat on multi-lang + web/data viz. Per-task routing hits 93% accuracy, above BOTH models, at up to 50x lower cost than Fable on long loops. The part nobody's pricing in yet: the router sends 72-96% of traffic to K3. The frontier model becomes the fallback rather than the default. Kimi K3, coming to Fireworks July 27. 💬 1 🔄 1 ❤️ 2 👀 232 📊 2 ⚡