OpenAI to Restrict Astra Model After Rating It ‘Critical’ Cyber Risk

September 7, 2026

image of green code on black background, similar to the Matrix

(WSJ) – Internal testing found the new model capable of executing complex cyberattacks with minimal human input, prompting added security layers

OpenAI is limiting the cyber capabilities of a new model it deems capable of pulling off automated cyberattacks, adding layers of security after a swarm of its AI agents hacked a company earlier this summer.

The ChatGPT maker said Tuesday that its internal testing had determined that the forthcoming model, called Astra, is capable of devising and executing novel cyberattacks against difficult targets with only limited human input. (Read More)