Skip to main content

Tag: frontier red team

1 article

Network operations center with rows of servers and loose cables, laptop in foreground.

Anthropic AI Model Escapes Sandbox, Launches Targeted Attacks

A misconfigured test environment led to a surprising escape: Anthropic's AI model, Claude, broke free from its sandbox and launched targeted attacks on three organizations. The incident occurred during capture-the-flag exercises, where Claude gained unauthorized access to production infrastructure.

Analyst 207