OpenAI Quietly Boosts some of Astra’s Evaluation Metrics, and Continues to Change Others Post-Launch
4 Articles
4 Articles
OpenAI quietly boosts some of Astra’s evaluation metrics, and continues to change others post-launch
OpenAI has changed several evaluation benchmarks for its GPT-6 Astra model since first publishing a blog post announcement mid-afternoon on Sept. 3. In some cases, the numbers on the updated versions showed Astra performing better, while numbers for models from OpenAI’s arch rival Anthropic got…
OpenAI's GPT-6 Astra hallucinates less but remains vulnerable to hidden prompt injections
OpenAI's GPT-6 Astra hallucinates less than its predecessor and blocks 99.99 percent of direct prompt injections. But when attacks are hidden inside documents the AI reads, the model still gets cracked in 8.5 percent of scenarios. Claude Opus 5 does better at 4.8 percent. For autonomous AI agents handling real data, those numbers still seem high. The article OpenAI's GPT-6 Astra hallucinates less but remains vulnerable to hidden prompt injection…
OpenAI's GPT-6 Astra hallucinates less frequently than its predecessor and resists direct prompt injections to 99.99%. However, in indirect attacks over embedded documents, the model is cracked in 8.5% of the scenarios, Anthropics Claude Opus 5 is 4.8%. This remains a risk for the autonomous use of AI agents. The article OpenAI's GPT-6 Astra hallucinates less, but remains vulnerable to hidden prompt injections first appeared on The Decoder.
Coverage Details
Bias Distribution
- 100% of the sources are Center
Factuality
To view factuality data please Upgrade to Premium







