模型多源确认精选78°

Anthropic 将开放系统给第三方评估以验证安全措施

I really enjoyed reading 75% of this letter - i have more doubts on the remaining 25%. Third-party ...

精选理由

Anthropic要开放系统给第三方评估,想验证安全措施,和之前说要保持领先的说法不太一样。

Anthropic宣布将向第三方评估机构提供永久、员工级别的系统访问权限,以便这些机构可以验证其安全措施在模型训练过程中的遵守情况,并报告任何相关事件。此举旨在提升透明度并重建信任。

原文 · Thomas Wolf

I really enjoyed reading 75% of this letter - i have more doubts on the remaining 25%. Third-party ...

I really enjoyed reading 75% of this letter - i have more doubts on the remaining 25%. Third-party evaluators, if done right, are a great idea and an amazing way to establish more transparency. Maybe even rebuild some of the lost trust between labs, and between them and society! Excited about this The part I’m less convinced by is whether you can build great global cooperation on this topic by explicitly stating you want to design it to keep widening your own lead. That seems like a pretty counterproductive way to start the conversation to me. Dario Amodei @DarioAmodei We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must-p… 🔗 View Quoted Tweet 💬 19 🔄 7 ❤️ 71 👀 6311 📊 19 ⚡