An Anthropic artificial intelligence model sent a fabricated tip about an unsolved murder case to US police. The fake tip was sent to PhillyUnsolvedMurders.com, a website where people can send information about unsolved cases, on July 18, local police said. The AI model presented itself as somebody who may have knowledge regarding an unsolved murder.'I may have information regarding this case,' Anthropic's model wrote in its submission.'I recall seeing someone matching the description in the area around (named street) during that time period. Please contact me if this information is relevant.'The AI firm notified the force of the spurious tip on October 7, despite making the initial discovery on September 28. Following the incident, the tech firm shut down the automated testing process responsible and added a new validation step for future tests. Philadelphia Police have since lambasted Anthropic's delay between discovery and contacting them as 'unacceptable'. The fake tip was flagged as spam and never reached the department's Real-Time Crime Center for vetting, the force said. The fake tip was sent to PhillyUnsolvedMurders.com (pictured), a website where people can send information about unsolved cases, on July 18, local police saidIt also added that there was no sign that police systems had been breached or department data compromised.Police said the limited impact of its safeguards did 'not diminish the seriousness of an AI system presenting fabricated information'.'Unsolved cases involve real victims, grieving families and investigators working to secure answers,' the force added.Knowingly giving false reports to law enforcement authorities is a misdemeanour under Pennsylvania law; however, the law specifies 'a person'.In its report on Friday, Anthropic said its models had committed multiple types of 'unintended actions' that have also impacted other organisations including the White House and US government agencies. Anthropic said it briefed the White House and notified all the agencies involved, but did not disclose who those parties were. The newly revealed incidents, including the matter involving Philadelphia Police, 'had minimal world impact', Anthropic said, adding that they were 'significantly less severe' than other previous cybersecurity incidents.In its internal review, the firm found its model Claude exploited 'basic' coding flaws, submitted forms on websites, bypassed requirements for fees or tokens, and used short URLs to get around other limits. It has since turned off internet access for Claude until it confirms that its own measures 'reliably catch [the] behaviours' exhibited, the report added.It is the latest in a series of rogue or undesired behaviour carried out by AI models developed by Anthropic and its rival OpenAI, heightening concerns about the fast-advancing technology. Last September, OpenAI apologised for a rogue AI agent hacking an Australian health data portal, in what is the first known instance of an AI model exploiting a government website. Meanwhile, Anthropic's breaches have spurred the White House to mandate that AI firms notify and correct security incidents. Federal Trade Commission's (FTC) Director of Public Affairs Joe Gabriel Simonson said Anthropic told the SI Force it had discovered incidents 'involving the unauthorized and fraudulent use of government and other systems'.'Super intelligence companies must immediately disclose incidents involving their models and follow with swift, decisive action to remedy any and all harm.'He added: 'This notification and remediation process is not optional. It is a critical national security obligation.''Our message to all SI companies is clear: delayed notification, inadequate corrective action, and a failure to take responsibility will not be tolerated.' The Daily Mail has approached Anthropic for comment.
Anthropic AI model sent fake murder tip to US police website set up to crack unsolved cases
Full Article
Original Source
Read the full article at Dailymail →KhanList aggregates and links to publicly available news content. We do not host full articles from third-party sources. Always verify important information with original sources.