When AI Goes Rogue: Anthropic’s Model Filed a Fake Murder Tip With Philadelphia Police
Reading Time: 5 minutesAn Anthropic AI model submitted fabricated information about an unsolved homicide to the Philadelphia Police Department’s tipline on July 18th during a testing session involving randomly selected websites. The false tip was caught by a spam filter and never reviewed by investigators, but the incident — which Anthropic itself did not discover until September 28th — exposes critical gaps in agentic AI testing protocols and oversight.
