GPT-6 Astra Claims AGI, But Hidden Documents Still Fool It #
OpenAI has released GPT-6 Astra to Pro, Enterprise, and Business Premium users, while President Greg Brockman calls it the start of the “AGI era.” The model blocks 99.99 percent of direct prompt injections, but hidden attacks inside documents it reads still crack it in 8.5 percent of scenarios, compared with 4.8 percent for Claude Opus 5.
Astra handles math, coding, cybersecurity, computer use, and browser tasks, according to OpenAI and early reports. Its standard-model message allowance is roughly half that of GPT-5.6 Sol, and Plus users are expected to get access soon. In an OpenAI case study, legal technology company Legora reviewed 41 documents in minutes, found all four planted errors, and improved performance by nearly 40 percent.
Next Big Future reported Astra at 98.6 percent on ARC-AGI-3 versus 30.2 percent for Claude Opus 5, and at 97.6 percent on FrontierMath Tier 4 versus 87.8 percent for Claude Fable 5.1. Those benchmark results sit alongside Brockman’s separate claim that Astra marks the beginning of the “AGI era”; the document-injection tests show that the model remains vulnerable when hidden instructions are embedded in files.
Why it matters: OpenAI gets a reported lead over Anthropic’s Claude models, but enterprise teams that feed Astra documents must plan around hidden instructions fooling it in 8.5 percent of tested scenarios while message allowances run at roughly half GPT-5.6 Sol’s.
Key Takeaways
- Pro, Enterprise, and Business Premium users get Astra first; Plus users are expected to follow soon.
- Legora reviewed 41 documents in minutes, caught four planted errors, and improved performance by nearly 40 percent.
- Astra scored 97.6 percent on FrontierMath Tier 4, versus 87.8 percent for Claude Fable 5.1 and 73.2 percent for Claude Opus 5.