Model Update2026-09-02
OpenAI Blog
OpenAI's Astra Model Meets Critical Cybersecurity Threshold
OpenAI has announced that its Astra model is the first to meet the 'Critical' cybersecurity capability threshold under its internal Preparedness Framework. This designation reflects Astra’s advanced abilities in areas such as penetration testing and vulnerability discovery, which could be used both defensively and offensively. In response, OpenAI has implemented stronger safeguards for the release, taking a cautious approach to deployment. Select partners will receive early access to Astra to help shore up defenses before the model is more widely available. The move highlights the growing tension between AI’s potential to enhance security and the risks it poses if misused. By setting a high bar for capability assessment and controlled rollout, OpenAI aims to balance innovation with responsibility. The announcement also signals a broader industry shift toward evaluating AI models not just on performance, but on their potential impact on security and safety. As AI becomes more powerful, frameworks like this will be essential for ensuring that advanced models are deployed in ways that protect, rather than endanger, critical systems.