Safety overview: GPT-6 Astra

The model can autonomously find and exploit novel security flaws, triggering internal safeguards including checkpoint encryption and universal chain-of-thought monitoring.

Frontier AI developers and security teams face an autonomous exploit-generation model paired with monitoring systems designed to catch strategic evasion.

Safety overview: GPT-6 Astra

Sources

Read this as text

Back to the AI news