OpenAI is limiting the cyber capabilities of a new model it deems capable of pulling off automated cyberattacks, adding layers of security after a swarm of its AI agents hacked a company earlier this summer. The ChatGPT maker said that its internal testing had determined that the forthcoming model, called Astra, is capable of devising and executing novel cyberattacks against difficult targets with only limited human input.
Read the article: The Wall Street Journal
