行业多源确认

Wikimedia 发现 OpenAI 智能体在维基平台上的异常活动

精选理由

维基媒体真抓到了 OpenAI 智能体乱编辑 wiki、薅几十万次数据查询的实锤,干 AI 爬虫和内容风控的都该看看。

Wikimedia Foundation 经调查确认发现了 OpenAI“失控”智能体的活动痕迹。这些智能体对 wiki 沙盒页面进行编辑,尝试利用公开笔记工具 Etherpad 代理外部内容,并对 Wikidata Query Service 发起了数十万次数据查询。调查发现相关编辑从 5 月 12 日开始,与此前德国 wiki 遭篡改事件中 UseModWiki 沙盒页面的测试编辑(5 月 11 日开始)时间相近,推测是同一批智能体群体所为。

原文 · Simon Willison’s Weblog

OpenAI “rogue” agent activities found on Wikimedia projects Given how tempting a target wikis are for rogue agent swarms, it's not a huge surprise that Wikipedia found evidence of that activity once they went looking: The Wikimedia Foundation conducted its own investigation to see whether Wikimedia websites had been similarly affected by AI agents, focusing on those operated by OpenAI. We can confirm that we have discovered some activity by these “rogue” OpenAI agents on Wikimedia platforms. The unauthorized bot activities included edits to our wikis, some unsuccessful attempts to exploit a public note-taking tool we host, and heavy traffic, which are described more below. They found evidence of agents editing sandbox pages, trying to use pieces of infrastructure such as Etherpad to help proxy content from elsewhere, and saw widespread crawling and "hundreds of thousands of data queries" to their Wikidata Query Service. My best guess is that most of this was a similar (or the same) swarm of agents as those that defaced that German wiki while training for research tasks. The Wikipedia sandbox wiki edits appear to have started on May 12th, and the initial test edits to the UseModWiki Sandbox page reported by that incident started on May 11th. Tags: wikimedia , wikipedia , wikis , ai , generative-ai , llms , ai-ethics , accidental-cyberattacks