OpenAI Releases GPT-6 Astra: A Model Capable of Autonomous Hacking
OpenAI has released GPT-6 Astra, a model that its president Greg Brockman calls a 'generational leap in capability' and the arrival of artificial general intelligence (AGI).
Astra is the first system OpenAI has rated capable of autonomously hacking well-protected systems without human guidance, raising safety and security concerns.
The model scored 100% on ExploitBench, a benchmark that measures a model's ability to turn known software flaws into functioning attacks. It also discovered two previously unknown zero-day vulnerabilities in Google's V8 engine during testing.
Astra is OpenAI's first model to cross the 'critical' threshold under its Preparedness Framework, the company's internal scoring system for dangerous capabilities.