Agent Playground发布,比较Claude Code、Codex等模型成本

Very cool launch from the @agentsky_dev team. Agent Playground lets you give Claude Code, Codex, and...

精选理由

AgentSky.dev发布,提供统一API接入多个云上智能体,并支持在浏览器中比较不同模型的表现,成本仅为2美元,值得尝试。

AI 摘要

Agent Playground允许用户在浏览器中比较Claude Code、Codex和DeepSeek等模型在相同任务下的时间、成本和token消耗。AgentSky.dev提供统一API接入多个云上主要智能体,并支持在浏览器中比较不同模型的表现。测试显示,使用AgentSky.dev的模型成本仅为2美元,而另一模型成本为150美元。

原文 · elvis

Very cool launch from the @agentsky_dev team. Agent Playground lets you give Claude Code, Codex, and...

Very cool launch from the @agentsky_dev team. Agent Playground lets you give Claude Code, Codex, and DeepSeek the identical task in one browser and compare time, cost, and tokens side by side. I run same-task harness tests constantly, and I think this is one of the first places agents can be compared under truly identical conditions. It's great because you can make a better decision about which agent harness is best for the desired task. Xiaoyin Qu @quxiaoyin I ran the same task on Claude Code and DeepSeek's new agent harness. One cost $150. The other cost $2. Today we're launching AgentSky.dev ( @agentsky_dev ), the "OpenRouter for Agents" — one API → Claude Code, Codex, DeepSeek, Kimi, OpenCode, and every major agent in the cloud. And Agent Playground on top: race them on your own task, with your real tools (GitHub, Gmail, more), side by side in a browser: time, cost, tokens burnt. Guess which one was $2. Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 0 🔄 0 ❤️ 4 👀 983 📊 1 ⚡