OpenRouter推出Ori Eval:用项目实测给模型打分选型

有点意思,别人选模型靠运气,Ori Eval 帮你靠考试,把模型拉到你的项目里实际做一遍题,谁分高用谁。

精选理由

OpenRouter出了个Ori Eval,能拿你项目的真实任务给模型打分,谁分高用谁,选模型不用瞎猜了。

AI 摘要

OpenRouter发布Ori Eval,用于在真实项目中评估模型表现。Ori Eval通过OpenRouter API对代码库各任务运行测试并给出评分。用户可通过curl命令获取该工具,例如从openrouter.ai下载。Ori Eval强调没有绝对最好的模型,只有最适合某个任务的模型。

原文 · Geek

有点意思,别人选模型靠运气,Ori Eval 帮你靠考试,把模型拉到你的项目里实际做一遍题,谁分高用谁。

有点意思,别人选模型靠运气,Ori Eval 帮你靠考试,把模型拉到你的项目里实际做一遍题,谁分高用谁。 OpenRouter @OpenRouter Introducing Ori Eval: the easiest way to write your first eval. There's no definitive best model, only the best model for each task. Ori Eval leverages OpenRouter's APIs for each task in your codebase, and then evaluates the results. curl -fsSL openrouter.ai/skills/spawn-o… 🔗 View Quoted Tweet 💬 0 🔄 0 ❤️ 1 👀 373 ⚡