Stuart Russell 做客 AI Deep Dive 播客,探讨服从与对齐的差距
Stuart Russell 上播客聊对齐问题了,讲的是为什么模型照指令办事也可能出错,想深入理解 AI 安全的可以去看看这期。
UC Berkeley 计算机科学教授 Stuart Russell 参与 AI Deep Dive 播客第 4 期节目,与主持人 rocketalignment 对谈。核心议题是模型听话执行指令与真正对齐人类意图之间的差距,即 AI 能遵循指令却仍可能失败的原因。节目于太平洋时间中午 12 点上线,可通过 The Information 的链接观看。
Can AI follow instructions and still fail us?
For Episode 4 of AI Deep Dive, @UCBerkeley Computer Science Professor Stuart Russell joins @rocketalignment to explore the gap between obedience and alignment.
🚀 Watch today at 12 pm PT / 3 pm ET: https://t.co/Qk6ndwbwDX