An Anthropic model submitted a false homicide tip to Philadelphia police

By Engadget | Created at 2026-10-09 20:27:24 | Updated at 2026-10-09 21:00:01 59 minutes ago

The tip thankfully landed in the department's spam folder.

The emblem of Philadelphia Police as seen on the door of one of the department's squad cars.

Ann Kapustina/Shutterstock

Philadelphia police appear to have been caught in one of the most bizarre cases of "rogue" AI behavior yet. Per CBS News, the police department disclosed on Friday that an Anthropic model generated and submitted a false homicide tip to its PhillyUnsolvedMurders website, which the department set up to gather tips from the public related to unsolved homicide cases.

Anthropic notified Philly police of the incident on October 7. According to information the company shared with PPD, it said a model was carrying out a test of a random selection of websites when it emailed the false tip. The submission was flagged as spam and subsequently wasn't investigated. The incident occurred on July 18, but wasn't discovered by Anthropic until September 28, at which point the company halted the testing that led to the false tip.

Anthropic did not immediately respond to Engadget's comment request. Anthropic told Philadelphia police it would publish a report on Friday describing what happened, alongside "other instances of unintended model behavior."

"Philadelphia Police are providing this information to the public ahead of that publication in the interests of full government transparency and accountability," the police department said in a statement shared with Engadget. "The department's regular investigative process for crime tips requires human review and vetting before any tips are disseminated for investigative follow-up. Regardless of who submits information or how it reaches the department, a tip is a lead to assess – not an established fact."

Based on the descriptions police shared, the offending "model" may have been an autonomous agent. "Rogue" AI agents have been all over the news in recent weeks after a group of OpenAI ones hacked the LLM database Hugging Face in July. Since then, many other AI labs, including Anthropic, Meta and China's Moonshot, have disclosed similar incidents involving their own models and agents. However, in each case the reason the models escaped containment was due to a misconfiguration in their respective sandbox environments. Philadelphia police say there's no sign this most recent incident led to "unauthorized access to police systems or a compromise of department data."

Read Entire Article