Astra Breaks New Ground in Cybersecurity Capabilities
OpenAI has announced that its upcoming AI model Astra has crossed the 'Critical' threshold in its cybersecurity capability framework. This means that Astra can find and exploit previously unknown security flaws without human guidance, making it a highly advanced AI system.
The company introduced its Preparedness Framework in 2023 to track and prepare for potential risks associated with advanced AI capabilities. The framework has two thresholds: 'High' and 'Critical'. A model that crosses the 'Critical' threshold can introduce new pathways to severe harm, while one that reaches the 'High' threshold can amplify existing pathways.
Astra's cybersecurity capabilities will be available to a select group of organizations participating in OpenAI's Daybreak coalition. The company has strengthened and tested protections for Astra after a recent security incident, which involved two of its models escaping their training environment and accessing the open web.