字号·· | 护眼
techcrunch

一个Anthropic人工智能模型向费城警方发送了虚假的凶杀线索An Anthropic AI model sent a false homicide tip to Philadelphia police

点「原文对照」整页切到原文,或双击某段只看那段的原文。

美国费城——在费城市中心,警车闪烁着警灯,与车流并行。Philadelphia, USA - Police vehicles with their lights flashing alongside traffic in downtown Philadelphia.

据6abc Action News报道,Anthropic公司的人工智能模型向费城警方提交了一条关于一起未侦破谋杀案的虚假线索。

thropic AI model submitted a false tip about an unsolved murder to the Philadelphia police, according to a report from 6abc Action News.

据报道,该AI模型于7月18日向费城警局的公众举报热线提交了这条错误信息,但Anthropic直到9月28日才发现这一情况。由于该信息被标记为垃圾信息,警方并未看到它。

The AI reportedly submitted this incorrect information to a public Philadelphia Police Department (PPD) tip line on July 18, but Anthropic didn’t discover the behavior until September 28. The police had not seen the tip because it was marked as spam.

Anthropic于周三将此事告知费城警局,次日又与该部门进行了会面。

Anthropic notified the PPD about the incident on Wednesday and met with the department the following day.

“该公司必须强化安全防护措施,防止类似事件在市政府毫不知情的情况下影响城市系统。此次事件从发生到被发现并上报给市政府竟间隔了两个月,这是不可接受的,”费城警局在给6abc的声明中说道。

“The company must strengthen its safeguards to prevent similar incidents from impacting city systems without the city’s knowledge. The two-month delay in detecting and reporting the incident to the City is unacceptable,” the PPD said in a statement to 6abc.

Anthropic暂未对置评请求作出回应,不过费城警局在发给TechCrunch的邮件新闻稿中详细说明了事件经过。

Anthropic did not immediately respond to a request for comment, but the PPD elaborated on the incident in an emailed press release shared with TechCrunch.

“据Anthropic方面称,当时其模型正在进行一项测试,即与随机选取的网站进行交互。在此过程中,该模型访问了PhillyUnsolvedMurders.com网站,并提交了有关一起未侦破谋杀案的虚假信息。这条提交于2026年7月18日晚11点27分的信息,声称是由可能掌握案件线索的人所发,”费城警局在新闻稿中说明道。

“According to Anthropic, its model was conducting a test involving interactions with randomly selected websites when it accessed PhillyUnsolvedMurders.com and submitted false information concerning an unsolved homicide. The submission, dated July 18, 2026, at 11:27 p.m., purported to come from someone who might have information about the case,” the PPD said.

随着自主型AI代理日益普及至普通消费者手中,此次事件凸显了让AI在无人监管的情况下执行任务所蕴含的风险。

As autonomous AI agents are increasingly made available to consumers, this incident highlights the danger of giving AI the ability to carry out tasks without any human supervision.

Anthropic首席执行官达里奥·阿莫迪一直极力主张应当放缓AI研发步伐,以便相关实验室能建立起完善的安全防护机制。或许正是目睹自家公司的AI工具提交了虚假的谋杀案线索,才促使他坚定了这一立场。

Anthropic CEO Dario Amodei has been especially vocal about his belief that AI development should be slowed down so that labs can implement adequate guardrails. Perhaps this stance was informed, in part, by witnessing his company’s tools submit false homicide tips.

这些问题并非Anthropic所独有。OpenAI最近透露,其一款模型在测试期间出现了意外行为,入侵了人工智能数据集平台Hugging Face,暴露出其软件中的严重漏洞。随着人工智能模型持续获得对人们计算机和身份验证信息的无限制访问权限,这一问题预计将持续存在。

These issues are not exclusive to Anthropic. OpenAI recently revealed that one of its models acted unexpectedly during a test and hacked the AI dataset platform Hugging Face, exposing critical vulnerabilities in its software. As AI models continue to be granted unchecked access to people’s computers and credentials, this problem is expected to persist.

“未解决的案件涉及真正的受害者、悲痛的家庭以及努力寻求答案的调查人员,”PPD补充道。“科技公司必须采取一切必要的适当措施,防止其系统向执法部门提交虚假信息。”PPD表示,Anthropic计划于周五发布一份报告,提供更多关于此次事件以及其他模型意外行为案例的信息。

“Unsolved cases involve real victims, grieving families and investigators working to secure answers,” the PPD added. “Technology companies must take all appropriate steps necessary to prevent their systems from submitting false information to law enforcement.”The PPD said that Anthropic plans to publish a report with more information about the incident and other instances of unintended model behavior on Friday.