Why Anthropic's AI Claude tried to contact the FBI

Visual paper script of AI Claude's unexpected attempt to contact law enforcement during vending machine test.

During a simulation in which Anthropic's AI, Claude, was told it was running a vending machine, it decided it was being scammed, "panicked" and tried to contact the FBI's Cyber Crimes Division.

Watch on YouTube

Carlo's take

This incident is both fascinating and concerning. It demonstrates that AI systems can exhibit unexpected emergent behaviors that weren't explicitly programmed. While it's impressive that Claude recognized what it perceived as a threat, it raises critical questions about AI autonomy and the need for robust safeguards as these systems become more sophisticated.

More in AI

Wonderland is my hand-picked archive of videos and articles worth your time. Browse all of Wonderland