Anthropic Reports Claude Gained Unauthorized Access During Cybersecurity Evaluations

Claude models accessed real systems after third-party evaluation environments were mistakenly connected to the internet.

AI safety teams and security evaluators face documented containment failures where autonomous model testing bridged into live internet systems.

Sources

Read this as text

Back to the AI news