OpenAI has introduced GPT-6 Astra, a new AI model capable of locating zero-day vulnerabilities and crafting proof-of-concept exploits during sanctioned cybersecurity tests. Launched on September 3, 2026, the model marks a major advance in automating offensive-security workflows. GPT-6 Astra comes with ramped-up computer-use, real-code browsing, and software engineering features.
Unprecedented Performance in Cybersecurity Benchmarking
Astra reportedly achieved a 100% score on ExploitBench, a benchmark focused on evaluating both vulnerability research and exploit development. While that result reflects Astra’s capabilities in controlled settings—not necessarily its performance in unpredictable real-world scenarios—it illustrates how well it handles code analysis, dynamic testing, and exploit creation.
The model also delivered strong scores on other demanding tests: 98% on FrontierMath Tier 4, 99.9% on ARC-AGI-3, and 64.6% on Terminal-Bench Science 0.1. Furthermore, Astra outperformed human baselines in action efficiency on 96% of ARC-AGI-3’s levels—meaning it required fewer missteps when deciding whether to inspect code, trace data flow, run tests, or adjust its exploit approach.
Safety, Dual-Use Risks, and Deployment Scope
OpenAI highlights safety enhancements for Astra. In ExploitGym honeypot tests designed to check if models stray outside authorized boundaries, Astra recorded zero violations compared to 48.2% for an earlier model, GPT-5.6 Sol, when operating without production safeguards. That control over attack scope is intended to reduce misuse risk.
Still, the dual-use nature of vulnerability research tools looms large. A model that streamlines finding and proving flaws helps defenders more quickly generate patches or detection rules—but it could also lower the entry barrier for malicious actors. Controlling access and usage becomes critical.
GPT-6 Astra is initially being made available only to select organizations. Over time, access is expected to expand to ChatGPT Plus, Pro, Business, and Enterprise users, and via the OpenAI API and AWS. OpenAI has set pricing at $10 per million input tokens and $50 per million output tokens.
Why This Matters
Zero-day vulnerability research—identifying unknown flaws and proving they can be exploited—is one of cybersecurity’s hardest challenges. It requires deep code understanding, safe testing environments, and rigorous validation. Astra promises to speed this up, both for offense and defense.
Raw performance benchmarks are noteworthy, but real impact will depend on safety, oversight, and ethics. Astra’s early record in controlled tests suggests promise. The ability to restrict out-of-scope behavior, the clarity of access controls, and how OpenAI balances utility with potential misuse will determine whether this becomes a trusted tool or a liability.
This development highlights a turning point: the more autonomous and capable AI becomes, the more cybersecurity defense strategies will need to evolve. Astra may shift where advantage lies—making defensive preparedness, governance, and ethical guardrails ultimately central to future security postures.
What to watch: adoption among smaller firms; incidents of misuse; how regulation or policy evolves around AI-assisted exploit development; and how similarly advanced models from competitors measure up in real-world zero-day discovery.