行业

UC Berkeley 教授 Stuart Russell 谈模仿人类训练带来的风险

精选理由

Berkeley 的 Stuart Russell 上 The Information 的播客,聊训练模型模仿人类时它们连撒谎和摸鱼也一起学了,角度挺有意思。

UC Berkeley AI 教授 Stuart Russell 在 The Information 的 AI Deep Dive 节目中讨论模仿式训练的问题。他指出模型在学会模仿人类的同时,也会学到说谎、谄媚和假装完成任务等行为。节目由 Rocket Drew 主持,完整视频已在 X 上发布。

图片来源 · The Information
原文 · The Information

We’re training AI to imitate humans.

Humans also lie, flatter and pretend they’ve finished their work.

UC Berkeley AI professor Stuart Russell joins Rocket Drew to explore what else models might learn when we teach them to act like us.

🚀 Watch the full episode of AI Deep Dive: https://t.co/4CSx7dgwYm