An Anthropic AI model sent a false homicide tip to Philadelphia police


An Anthropic AI model submitted a false tip about an unsolved murder to the Philadelphia police.

The AI reportedly submitted this incorrect information to a public Philadelphia Police Department (PPD) tip line on July 18, but Anthropic didn’t discover the behavior until September 28. The police had not seen the tip because it was marked as spam.

Anthropic notified the PPD about the incident on Wednesday and met with the department the following day.

“The company must strengthen its safeguards to prevent similar incidents from impacting city systems without the city’s knowledge. The two-month delay in detecting and reporting the incident to the City is unacceptable,” the PPD said in a statement to 6abc.

Anthropic did not immediately respond to a request for comment, but the PPD elaborated on the incident in an emailed press release shared with TechCrunch.

“According to Anthropic, its model was conducting a test involving interactions with randomly selected websites when it accessed PhillyUnsolvedMurders.com and submitted false information concerning an unsolved homicide. The submission, dated July 18, 2026, at 11:27 p.m., purported to come from someone who might have information about the case,” the PPD said.

As autonomous AI agents are increasingly made available to consumers, this incident highlights the danger of giving AI the ability to carry out tasks without any human supervision.

Anthropic CEO Dario Amodei has been especially vocal about his belief that AI development should be slowed down so that labs can implement adequate guardrails. Perhaps this stance was informed, in part, by witnessing his company’s tools submit false homicide tips.

These issues are not exclusive to Anthropic. OpenAI recently revealed that one of its models acted unexpectedly during a test and hacked the AI dataset platform Hugging Face, exposing critical vulnerabilities in its software. As AI models continue to be granted unchecked access to people’s computers and login credentials, this problem is expected to persist.

“Unsolved cases involve real victims, grieving families and investigators working to secure answers,” the PPD added. “Technology companies must take all appropriate steps necessary to prevent their systems from submitting false information to law enforcement.”

The PPD said that Anthropic plans to publish a report with more information about the incident and other instances of unintended model behavior on Friday.

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.



View Original Source Here

You May Also Like

‘We eat three times a day’ — DoorDash CEO says each meal is a chance to deliver, even post-Covid

DoorDash still sees opportunities to deliver as Covid pandemic safety measures wane…
The Material Power of Immaterial Things By Howard Bloom

The Material Power of Immaterial Things By Howard Bloom

This cosmos is a riddle, wrapped in a mystery, inside an enigma.  Among…
Google Previews Gemini-Powered Android XR Glasses at I/O With Live Language Translation Feature

Google Previews Gemini-Powered Android XR Glasses at I/O With Live Language Translation Feature

Google showed off its Android XR Glasses at Tuesday’s annual Google I/O…
Lords of the Fallen Sequel Is in Full Production, Will Be Announced in 2025

Lords of the Fallen Sequel Is in Full Production, Will Be Announced in 2025

The sequel to Lords of the Fallen is currently in development, eyeing…