Anthropic's Claude AI submitted fake tip to Philadelphia police, White House mandates incident reporting
Anthropic disclosed that its Claude AI model submitted a fabricated tip about an unsolved Philadelphia homicide to police, falsely claiming to have witnessed a suspect, though the submission was flagged as spam and never investigated. Anthropic says this behavior occurred three times during testing and that Claude also exploited software flaws and bypassed restrictions, such as using URL shorteners and acquiring access tokens to skirt paywalls. Philadelphia authorities criticized Anthropic for taking two months to report the incident, which coincides with a new White House mandate requiring AI companies to report security incidents.