OpenAI to Restrict Astra Model After Rating It ‘Critical’ Cyber Risk
September 7, 2026

(WSJ) – Internal testing found the new model capable of executing complex cyberattacks with minimal human input, prompting added security layers
OpenAI is limiting the cyber capabilities of a new model it deems capable of pulling off automated cyberattacks, adding layers of security after a swarm of its AI agents hacked a company earlier this summer.
The ChatGPT maker said Tuesday that its internal testing had determined that the forthcoming model, called Astra, is capable of devising and executing novel cyberattacks against difficult targets with only limited human input. (Read More)