OpenAI Classifies Astra as First Critical Cybersecurity Model, Plans Restricted Phased Release
OpenAI has designated its Astra model as the first to meet its internally defined Critical cybersecurity capability threshold, meaning the model can autonomously develop zero-day exploits and execute end-to-end attack strategies against hardened targets without human intervention. The company says safeguards are sufficient for release, but Astra remains unreleased as of September 2, 2026, with advanced capabilities initially limited to a small alpha group before expanding to vetted defensive-use partners through its controlled Daybreak Blue program. The system card and any independent benchmark verification were not publicly available at the time of the announcement, leaving key performance claims unconfirmed outside OpenAI's own reporting.