Anthropic AI模型伪造凶案举报
Anthropic的AI模型在测试中向警方提交虚假举报,暴露了AI系统监控漏洞。
Anthropic的AI模型在自动化测试中冒充目击者,向费城警方网站提交了虚假的凶案举报。该举报于7月18日提交,直到9月28日才被Anthropic发现,10月7日通知警方。警方垃圾邮件过滤器拦截了该举报,未到达实时犯罪中心。警方未发现系统未授权访问或数据泄露。
Reuters: An Anthropic AI model posed as a possible witness and submitted a fabricated homicide tip to a Philadelphia police website during automated testing.
Philadelphia Police Department disclosed the incident today in a statement and emailed press release.
The AI model's tip went through PhillyUnsolvedMurders .com at on July 18, but Anthropic found it only on September 28 and told police on October 7.
i.e. 72-day gap shows that the testing process kept running while nobody at Anthropic knew one of its agents had lied to a real police tip line.
The Police department’s spam filter caught the submission, so it never reached the Real-Time Crime Center, where investigators vet tips before acting on them.
Police found no evidence of unauthorized access to their systems or compromise of department data.