OpenAI launches GPT-6 Astra with hacking risks in check

September 4, 2026:

OpenAI launches GPT-6 Astra with hacking risks in check

What you need to know

  • GPT-6 Astra is the first OpenAI model to hit the “critical” cybersecurity threat level, but deployment is moving forward with significant safeguards in place.
  • It is currently limited to defensive tasks like secure code review. Advanced capabilities like exploit creation are blocked for now, but trusted users in the Daybreak program will eventually get access to complex workflows.
  • Astra drastically outperforms GPT-5.6 Sol, hitting a perfect 100% on ExploitBench and 42.4% on ExploitGym.

OpenAI just released GPT-6 Astra, and it’s the company’s first model to reach the “critical” threat level in the company’s Preparedness Framework for cybersecurity. Despite these unprecedented security risks, OpenAI is moving ahead with the deployment.

That doesn’t mean GPT-6 Astra is being rolled out as an unrestricted hacking tool. The version rolling out now will do defensive work such as secure code review and patching, but will not entertain more advanced requests like creating proof-of-concept exploits, OpenAI says.

Source link