行业多源确认83°

OpenAI 被指将已发表的 AI 安全发现据为己有

精选理由

Gary Marcus 曝 OpenAI 把别人 2025 年就发表的自复制提示注入说成自家发现,连披露报告的首条引用都是那篇论文,挺尴尬的。

Gary Marcus 发帖指出,OpenAI 在报告中把一种攻击向量称为自己的新研究发现,但实际上此前已有公开研究。网友 @Actuallykeltan 指出,搜索 Self Replicating Prompt Injection 的第一条结果就是 2025 年该攻击的实验演示。该论文作者去年就向 OpenAI 披露过发现,且论文正是 OpenAI 披露报告的首条引用。争议焦点在于 OpenAI 是刻意误述,还是用 GPT 生成引用后未做核查。

原文 · Gary Marcus

🚨BREAKING (and I am not making this up): OpenAI claimed credit for discovering a new vector of attack that was actually previously known and cited in their own report. 🤦‍♂️ keltan @Actuallykeltan WTF? "A new research finding" Try googling Self Replicating Prompt Injection. The first result is an experimental demo of this from 2025! Oh, and authors of that paper disclosed their findings to OpenAI last year. The paper itself is the top citation in the OAI disclosure report. It seems OAI is either intentionally misrepresenting this as their own novel finding, or (and I do not want to believe this myself) had GPT generate a list of citations for them, then did not check the citations. 🔗 View Quoted Tweet 💬 4 🔄 4 ❤️ 16 👀 1759 📊 5 ⚡