行业73°

Anthropic 撤回 Claude Fable 5 秘密降级政策,公开道歉

good. now let's undo the nerf stuff as well

精选理由

Anthropic 的公开道歉和策略调整,为 AI 行业树立了透明度标杆,做 AI 安全或竞争分析的从业者值得关注这一转折。

AI 摘要

Anthropic 因在 Claude Fable 5 中秘密降低对竞争 AI 研究者的性能而遭到强烈反对。公司宣布将修改安全措施,使其对前沿大模型开发透明可见。Anthropic 承认做出了错误的权衡,并为此道歉。这一事件凸显了 AI 公司在竞争与安全之间的平衡难题。

原文 · elvis

good. now let's undo the nerf stuff as well

good. now let's undo the nerf stuff as well Max Zeff @ZeffMax NEW: Anthropic is walking back Claude Fable 5's policy to covertly degrade performance for competing AI researchers, after facing fierce backlash. “We’re changing Fable 5’s safeguards for frontier LLM development to make them visible,” Anthropic tells WIRED. “We made the wrong tradeoff and we apologize for not getting the balance right.” 🔗 View Quoted Tweet 💬 5 🔄 0 ❤️ 9 👀 1394 📊 4 ⚡