WSJ:三名被 OpenAI 解雇的安全研究员请求董事会暂停削弱 AI 可监控性的开发
三个被 OpenAI 开除的安全研究员上书董事会,说别再削弱人类盯着 AI 推理的能力,OpenAI 回应说是泄露敏感信息。两边说法不一样,事件细节可以看看。
据 WSJ 报道,三名因涉嫌不当行为被 OpenAI 解雇的安全研究员向公司董事会提交请求,要求停止任何进一步削弱人类监控 AI 推理能力的开发。其中 Tomek Korbak 和 Mikita Balesni 是 2025 年一篇论文的主要作者,该论文指出 chain-of-thought 监控有用但脆弱。OpenAI 方面回应称,解雇原因是他们不当处理敏感信息,与安全立场无关。
WSJ: Three safety researchers fired by OpenAI for alleged misconduct have asked its board to halt any development that further reduces humans’ ability to monitor AI reasoning.
Two of them, Tomek Korbak and Mikita Balesni, led the 2025 paper that warned chain-of-thought monitoring is useful but fragile.
OpenAI however says the dismissals concerned mishandled sensitive information.