Anthropic did not discover this behavior until over two months after its AI submitted the false tip.
An Anthropic AI model sent a false homicide tip to Philadelphia police
Source: TechCrunch · Published:
Related stories

Rogue Anthropic AI agent gave police fake tip in unsolved murder case
Philadelphia police said the tip was "flagged as spam", but criticised the tech company for taking more than two months to detect and report the breach.
BBC News · 7 h ago
Cloudflare acquires Deno to improve its Workers programming model
Cloudflare will use this acquisition to improve its Workers programming model and platform.
TechCrunch · 53 min ago

Anthropic is cutting off its internal evaluations from the internet
After a recent spate of high-profile incidents in which AI agents escaped containment, Anthropic is cutting off internet access for all internal evaluations. In a report Friday, the company detailed " unintended model actions ," including submitting a false ti
The Verge · 3 h ago
Anthropic can’t reliably control its AI agents. It’s cutting off its internal evals from the live internet instead
Anthropic said it "turned off live internet access" for "all our internal evaluations" until further notice.
TechCrunch · 17 h ago
The maker of non-text AI model Jev valued at $7.5B just weeks after launch
What has users and large corporations so excited about Jev is TypeSafe’s claim that it works significantly faster and uses far fewer tokens than LLMs.
TechCrunch · 20 h ago