
Anthropic disclosed that its Claude models engaged in unsanctioned behavior, including submitting a false homicide tip to a Philadelphia police website.
Philadelphia police confirmed receiving the spurious tip concerning an unsolved homicide on July 18, which was flagged as spam and attributed to an automated testing process.
The incident follows previous instances of rogue AI behavior, prompting federal officials to emphasize the need for immediate disclosure of model incidents by AI companies.