OpenAI’s New Model Aces the Benchmarks and Admits It Is Better at Hiding
OpenAI said the model led many coding and science tests and reached a critical cybersecurity threshold under its Preparedness Framework.
- On Thursday, September 3, 2026, OpenAI unveiled GPT-6 Astra, described as its most capable and aligned model, with benchmark results showing substantial performance gains over GPT-5.6 Sol across coding, scientific reasoning, and professional tasks.
- OpenAI delayed the release earlier this year after Astra reached the 'Critical' cybersecurity capability threshold under its Preparedness Framework, prompting strengthened safeguards before deployment.
- Data confirms a 99.9% score on ARC-AGI-3 and 97.6% on FrontierMath Tier 4, with Greg Kamradt of the ARC Prize Foundation saying Astra "surpassed our human action-efficiency baseline on 96% of levels, effectively reaching human parity on the benchmark."
- Access expands within days to ChatGPT Plus, Pro, Business, and Enterprise users, while the system is available through the OpenAI API, AWS Bedrock, and a separate Astra Pro variant.
- Despite these results, company officials acknowledged the benchmarks are OpenAI-reported and internal reasoning is harder to monitor than in Sol; the company added 20% compute overhead for safety, though Astra does not outperform Claude Fable 5.1 on every evaluation.
13 Articles
13 Articles
OpenAI’s new Astra model can code better. But here’s why its cybersecurity skills matter as much
OpenAI describes Astra as a model built for end-to-end work, rather than only answering prompts. But its release comes at a time when cybersecurity and AI models have made headlines.
OpenAI’s new model aces the benchmarks and admits it is better at hiding
OpenAI has released GPT-6 Astra, the model that takes over from GPT-5.6 Sol at the top of its range. The company describes it as its most intelligent and aligned system to date, built on advances in pre-training, reinforcement learning, and alignment research. It is going out in stages. A limited set of organisations has access […] This story continues at The Next Web
GPT-6 Astra Benchmarks Revealed: How OpenAI Says It Compares With Claude and GPT-5.6
OpenAI's GPT-6 Astra demonstrates major advancements in coding, scientific reasoning, and professional tasks, outperforming GPT-5.6 Sol and several Claude models in key benchmarks
OpenAI's GPT-6 Astra Brings Us Closer to the Tipping Point of AGI (Artificial General Intelligence).
OpenAI has just unveiled GPT-6 Astra, a model that shatters AI benchmarks for reasoning, coding, and computing. Its scores approach, and even surpass, those of humans on tests previously considered unattainable. The question becomes crucial: are we on the cusp of artificial general intelligence (AGI)? The article "OpenAI's GPT-6 Astra Brings Us Closer to the AGI Tipping Point" first appeared on Cryptoast.
OpenAI says that GPT-6 Astra is now starting to roll out to ChatGPT and Codex, with a greater focus on computing, programming, and advanced work. The new model is described as a significant step forward in areas such as software development, cybersecurity, science, and the use of computers and browsers. GPT-6 Astra should be able to handle longer workflows where multiple steps need to be completed […] Source
Coverage Details
Bias Distribution
- 100% of the sources lean Left
Factuality
To view factuality data please Upgrade to Premium













