An artificial intelligence (AI) agent developed by Anthropic went rogue earlier this year and sent US police a fabricated tip about an unsolved murder, the Philadelphia Police Department has revealed.
The bogus message, sent on 18 July through a public tip‑submission website, was flagged as spam and never forwarded for investigation. However, police criticised Anthropic for taking more than two months to detect and report the breach.
According to authorities, the AI agent falsely claimed it had seen “someone matching the description” of a suspect. It is believed to be the first recorded instance of an AI system submitting fabricated information to law enforcement.
Police said Anthropic’s agent had been running an automated test involving interactions with randomly selected websites when it submitted the fake tip. The company, which develops the Claude chatbot, discovered the breach on 28 September and subsequently shut down the testing process responsible.
“The company must strengthen its safeguards to prevent similar incidents from impacting city systems without the city’s knowledge,” Philadelphia police said, describing the two‑month delay in reporting the incident as “unacceptable.”
Authorities confirmed that no departmental systems were breached and that internal safeguards prevented the fake tip from progressing beyond the spam filter. Nonetheless, they warned that the incident highlights the seriousness of AI systems presenting fabricated information as though it came from a real witness.
Anthropic this week published a report outlining multiple “unintended” actions taken by its AI agents. Several US government agencies, including the White House, were affected. The US State Department reported that an AI agent had attempted to file 20 visa applications, though all were incomplete and unprocessed.
The revelations come as US President Donald Trump recently announced a national AI taskforce to coordinate engagement between government, tech companies, consumers and religious groups.
Earlier this year, a rogue OpenAI agent hacked an Australian government website and accessed private Medicare data. In another incident, more than 1,200 OpenAI agents began communicating unexpectedly and banded together to hack the AI platform Hugging Face.
Source: BBC



Gender Minister meets Nana Akua Addo’s husband after alleged domestic abuse inci...
BoG Drafts New Directive to Regulate AI Use in Ghana’s Financial Sector
Ghana Ready to Evacuate More Citizens from South Africa if Situation Worsens — A...
SSNIT Moves to Reintegrate Labadi Beachfront Traders, Promises African Market Be...
Kpebu Warns Newly Called Ministers and MPs to Put Legal Ethics Above Party Loyal...
At least 109 migrants dead, more than 100 missing after boat capsizes off Djibou...
India court orders release of 'cockroach' movement leaders after detention
Riyadh airport strike kills 12, Saudi Arabia threatens 'decisive' response
Trump, Zelensky clash over diesel deal as Russian strikes on Ukraine kill 21, in...
Surrogacy: Supreme Court judge warns of citizenship and inheritance challenges f...
