想省钱就用 Kimi K3,想快就用 Fable 5,两个都能修好 bug,实测数据给你参考。
Cline 团队在相同 CLI 环境下测试 Kimi K3 和 Claude Fable 5 修复同一仓库 bug。Fable 5 耗时 3.5 分钟、18 次工具调用,Kimi K3 耗时 12 分钟、34 次工具调用,慢 3.4 倍。Token 消耗上 Kimi 用 120 万,Fable 用 73 万,Kimi 多 1.7 倍。但费用 Kimi 仅 $0.92,Fable 需 $2.13,Kimi 便宜 2.3 倍,源于其每 token 定价 $3/$15 对比 Fable 的 $10/$50。这是首次开源模型在修复任务中与顶尖闭源模型正面竞争。
Kimi K3 和 Claude Fable 5 真实 BUG 修复对比! @cline 用自己仓库里的一个真实 bug,在完全相同的 Cline CLI harness 下让 Kimi K3 和 ...
Kimi K3 和 Claude Fable 5 真实 BUG 修复对比! @cline 用自己仓库里的一个真实 bug,在完全相同的 Cline CLI harness 下让 Kimi K3 和 Claude Fable 5 两个模型各跑一次修复任务。 结果两个模型都成功修复了 bug,差异在于过程效率: · 速度:Fable 5 用 3.5 分钟、18 次工具调用完成;Kimi K3 用了 12 分钟、34 次工具调用,慢 3.4 倍 · Token 消耗:Kimi 用了 120 万 token,是 Fable(73 万)的 1.7 倍 · 费用:Kimi 反而便宜 2.3 倍($0.92 vs $2.13),完全靠单价优势——Kimi 定价 $3/$15,Fable $10/$50,每 token 便宜约 3.3 倍 Cline @cline We tested Kimi K3 and Fable on a real bug from the Cline repo, and found that while both models were able to fix it - Fable wins on speed & Kimi wins on cost. - Kimi used 1.7x more tokens than Fable (1.2M vs. 730K) - Fable finished 3.4x faster - 3.5 min and 18 tool calls vs. Kimi’s 12 min and 34 tool calls. - Kimi cost 2.3x less ($0.92 vs. $2.13) thanks to its 3.3x per-token discount Both runs used the same Cline harness, and the traces indicate that Kimi is RL trained to spend more tokens thinking and verifying before completing. This is the first time we've seen an open weight model compete head to head with SOTA. Congratulations to the @Kimi_Moonshot team on this milestone! 🔗 View Quoted Tweet 💬 3 🔄 0 ❤️ 0 👀 320 📊 3 ⚡