daridotdev搞了个开源自路由模型,切换模型前先算缓存,真能帮你省钱还保持编码水平,值得试试。
daridotdev发布了开源权重自动路由模型,专为编码代理设计,在帕累托前沿达到SOTA,成本降低70%且编码性能与Fable相当。该模型采用缓存感知切换策略,避免因模型切换重置提示缓存导致的高费用,支持Claude Code、Codex等现有工具。
Switching models mid-session is one of the fastest ways to inflate an agent bill. Every switch res...
Switching models mid-session is one of the fastest ways to inflate an agent bill. Every switch resets your prompt cache, and a cheaper model can end up costing you more. @daridotdev built their router to be cache-aware. It only switches when the move actually saves money, and it plugs into Claude Code, Codex, or whatever harness you already use. The router itself is an open-weight SLM. Avyay Varadarajan @avyvar Today, we're releasing our open-weight, auto-routing model @daridotdev , built for coding agents. We're state-of-the-art on the Pareto Frontier, w/ 70% cost reduction + comparable coding performance to Fable. Bring your own evals, choose your models, or use our defaults. 🔗 View Quoted Tweet 💬 0 🔄 0 ❤️ 3 👀 748 📊 1 ⚡