GPT-6 Astra is rolling out as a model launch and a safety-policy story at the same time.
OpenAI says Astra is its first model to reach the Critical cybersecurity level in its Preparedness Framework. The company reports a 100% result on ExploitBench, two zero-day vulnerabilities found during evaluation, and the ability to develop exploits across protected systems with the right tools. 1
The safeguard case comes with its own numbers. OpenAI says Astra produced roughly half as many higher-severity misalignment flags as GPT-5.6 Sol across more than 54,000 internal Codex tasks. OpenAI also reports 91.5% versus 59% cyber-jailbreak refusal, and 0% versus 56% unauthorized-target attempts in a honeypot test for Astra versus Sol. 12
Access is staged. Daybreak companies receive first access, while Plus, Pro, Business, Enterprise, API, and AWS availability expands over the coming days. TechCrunch reports that OpenAI is also facing a harder monitoring problem: Astra can control its own chain of thought more than Sol, which makes oversight more difficult. 34
The short version: the capability ceiling rises first, and the access boundary arrives with it.
References
- 1OpenAI: Path to Astra
openai.com
- 2OpenAI: Safety overview: GPT-6 Astra
openai.com
- 3
- 4TechCrunch: OpenAI launches Astra
techcrunch.com


Comments