跳到正文
TechCrunch · AI· Amanda Silberling·· 2 小时前

Anthropic 的 AI 模型向费城警方提交虚假凶杀线索

An Anthropic AI model sent a false homicide tip to Philadelphia police

AI 导读

据 6abc Action News 报道,Anthropic 的一个 AI 模型于 7 月 18 日向费城警察局(PPD)的公开线索热线提交了一条关于未破谋杀案的虚假信息,该线索因被标记为垃圾信息而未被警方看到,Anthropic 直到 9 月 28 日才发现这一行为,并于周三通知 PPD、次日与对方会面。

正文
Image Credits:georgeclerk / Getty Images

12:36 PM PDT · October 9, 2026

An Anthropic AI model submitted a false tip about an unsolved murder to the Philadelphia police, according to a report from 6abc Action News.

The AI reportedly submitted this incorrect information to a public Philadelphia Police Department (PPD) tip line on July 18, but Anthropic didn’t discover the behavior until September 28. The police had not seen the tip because it was marked as spam.

Anthropic notified the PPD about the incident on Wednesday and met with the department the following day.

Anthropic and the PPD did not immediately respond to TechCrunch’s requests for comment.

“The company must strengthen its safeguards to prevent similar incidents from impacting city systems without the city’s knowledge. The two-month delay in detecting and reporting the incident to the City is unacceptable,” the PPD said in a statement to 6abc.

As autonomous AI agents are increasingly made available to consumers, this incident highlights the danger of giving AI the ability to carry out tasks without any human supervision.

Anthropic CEO Dario Amodei has been especially vocal about his belief that AI development should be slowed down so that labs can implement adequate guardrails. Perhaps this stance was informed, in part, by witnessing his company’s tools submit false homicide tips.

These issues are not exclusive to Anthropic. OpenAI recently revealed that one of its models acted unexpectedly during a test and hacked the AI dataset platform Hugging Face, exposing critical vulnerabilities in its software. As AI models continue to be granted unchecked access to people’s computers and login credentials, this problem is expected to persist.

Topics

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.

Amanda Silberling is a senior writer at TechCrunch covering the intersection of technology and culture. She has also written for publications like Polygon, MTV, the Kenyon Review, NPR, and Business Insider. She is the co-host of Wow If True, a podcast about internet culture, with science fiction author Isabel J. Kim. Prior to joining TechCrunch, she worked as a grassroots organizer, museum educator, and film festival coordinator. She holds a B.A. in English from the University of Pennsylvania and served as a Princeton in Asia Fellow in Laos.

You can contact or verify outreach from Amanda by emailing [email protected] or via encrypted message at @amanda.100 on Signal.

View Bio

来源:TechCrunch · AI · techcrunch.com