用户评测:GPT 5.6 Sol 表现出色,Cursor Grok 4.5 令人愉悦

My evaluation of some of the models over past ~24 hours testing/building: GPT 5.6 Sol is excellent....

精选理由

Suhail 实测了 GPT 5.6 Sol、muse-spark-1.1 和 Cursor Grok 4.5,GPT 5.6 Sol 和 Cursor Grok 4.5 表现抢眼,后者快速模式尤其好用。

AI 摘要

Suhail 分享了近24小时对多个模型的测试与构建体验。GPT 5.6 Sol 获得高度评价,认为其细节处理出色。muse-spark-1.1 在真实问题测试中表现不佳,存在格式差、深度不足、回复过于功利等问题。Cursor Grok 4.5 的快速模式使用体验愉悦,能深入理解代码库,逐渐赢得用户信任。

原文 · Suhail

My evaluation of some of the models over past ~24 hours testing/building: GPT 5.6 Sol is excellent....

My evaluation of some of the models over past ~24 hours testing/building: GPT 5.6 Sol is excellent. You can feel the care in it. Bravo as usual OpenAI. muse-spark-1.1 is way off the mark after my testing of some real questions. Bad formatting, not thorough, doesn't dig deeper, seemingly afraid to respond at length, too utilitarian. Something is off. Cursor Grok 4.5 has been joyful to use. I ask it lots of side questions while Claude codes. The fast mode is a real delight each day. It really understands, chomps through the code base, and has gotten me to trust it more each day. 💬 11 🔄 4 ❤️ 94 👀 9069 📊 21 ⚡