Skip to main content
See every side of every news story
Published loading...Updated

OpenAI Quietly Boosts some of Astra’s Evaluation Metrics, and Continues to Change Others Post-Launch

Summary by BizToc
OpenAI has changed several evaluation benchmarks for its GPT-6 Astra model since first publishing a blog post announcement mid-afternoon on Sept. 3. In some cases, the numbers on the updated versions showed Astra performing better, while numbers for models from OpenAI’s arch rival Anthropic got…

4 Articles

OpenAI's GPT-6 Astra hallucinates less frequently than its predecessor and resists direct prompt injections to 99.99%. However, in indirect attacks over embedded documents, the model is cracked in 8.5% of the scenarios, Anthropics Claude Opus 5 is 4.8%. This remains a risk for the autonomous use of AI agents. The article OpenAI's GPT-6 Astra hallucinates less, but remains vulnerable to hidden prompt injections first appeared on The Decoder.

·Germany
Read Full Article
Think freely.Subscribe and get full access to Ground NewsSubscriptions start at $9.99/yearSubscribe

Bias Distribution

  • 100% of the sources are Center
100% Center

Factuality Info Icon

To view factuality data please Upgrade to Premium

Ownership

Info Icon

To view ownership data please Upgrade to Vantage

Fortune broke the news in New York, United States on Friday, September 4, 2026.
Too Big Arrow Icon
Sources are mostly out of (0)

Similar News Topics

News
Feed Dots Icon
For You
Search Icon
Search
Blindspot LogoBlindspotLocal