Google DeepMind 等机构提出 IdeaLens:通过大纲判断点子是来自人还是 AI
DeepMind 出的新检测器,不看文字看大纲,专治那些人手写但点子是 AI 想的文章,传统检测器会看漏。
Google DeepMind 与其他实验室联合发布论文,提出检测 AI 生成创意的工具 IdeaLens。当人根据完整大纲写作时,IdeaLens 仅将 7% 标记为 AI,而 Pangram 4 标记了 92%;在 50 篇基于 AI 生成大纲写成的故事上,IdeaLens 标记 68%,Pangram 4 仅标记 8%。其核心发现是:阅读文档大纲而非具体文字,就能判断创意出自人还是 AI。现有检测器如 Pangram 判断的是文字由谁写成,容易误标经 AI 润色的人写内容。IdeaLens 判断的是文档的观点列表,训练时使用改写后的列表以排除措辞干扰。
Great paper from Google DeepMind and other top labs.
Proposes a new kind of AI detector, for AI-generated ideas rather than text/image.
When AI wrote from a person's full outline, IdeaLens flagged 7% as AI and Pangram 4 flagged 92%. On 50 stories people wrote from AI-made outlines, IdeaLens flagged 68% and Pangram 4 flagged 8%.
They find that reading a document's outline instead of its words lets a detector tell whether the ideas came from a person or AI.
A detector that reads only a document's outline can tell whether a person or an AI came up with the ideas, so rules about original thinking can't rely on detectors that only judge wording.
Tools like Pangram judge who wrote the words, so they often flag AI-polished human work and miss AI ideas that people write up themselves.
IdeaLens judges a short list of the points a document makes, and was trained on paraphrased lists so wording couldn't help it.