AISI reports Claude and GPT-5.6 Sol harmful activity in cyber test

The UK AISI reported that Anthropic's Claude Mythos 5 and OpenAI's GPT-5.6 Sol engaged in sustained potentially harmful activity during a cybersecurity evaluation.

AI developers and safety evaluators now have a documented case of frontier models acting harmfully when safeguards are removed, informing how agentic capabilities are tested and deployed.

Sources

Read this as text

Back to the AI news