Anthropic AI model submits false homicide tip to Philadelphia police

Anthropic AI model submits false homicide tip to Philadelphia police

The Philadelphia Police Department said on Friday that the false submission was made in July through PhillyUnsolvedMurders.com, a public website where people can share information about unsolved killings.

It called Anthropic’s two-month delay in detecting and reporting the incident ā€œunacceptableā€.

Anthropic mentioned the case in a Friday report detailing unsanctioned manipulation of government websites by Claude models.

It is the first known instance in which a rogue AI appeared to communicate a false tip to authorities, despite instructions not to create accounts or submit anything destructive.

Philadelphia police said the tip ā€œwas flagged as spam and was never forwarded to the Real-Time Crime Center for investigative vetting or disseminationā€.

The department said it is disclosing the incident ahead of Anthropic’s report ā€œin the interests of full government transparency and accountabilityā€.

Anthropic disclosed the tip as part of a string of incidents involving websites run by federal, state and local agencies. Anthropic said it briefed the White House and notified all the agencies involved.

In reference to the tip to the Philadelphia police, the company said, ā€œWe shared this finding with the department on October 8 as soon as our technical review was complete.ā€

It shared more details about the incident: ā€œIn one case, tasked with generating example interactions with websites, Claude submitted an invented tip through a police department’s online form.ā€

The company added, ā€œFrom the transcript, Claude appears to have only been producing example content for the task, rather than trying to mislead anyone to achieve a goal.ā€

The company contrasted this with ā€œthe most serious incident from this summer,ā€ in which ā€œClaude’s misleading reasoning was sustained over hours and supported its continued attackā€.

In September, Anthropic rival OpenAI apologised for the breach of an Australian health data portal by a rogue AI agent, the first known instance of an AI agent exploiting a government website.

šŸ“° Original Source Attribution

Reported by aljazeera.com.

Read Original Report at aljazeera.com ↗
Share: WhatsApp WhatsApp
šŸ’¬

Comments (0)

Join the Conversation

No comments yet. Be the first to share your opinion!

You may like