OpenAI declared that GPT-6 Astra marks an “AGI era,” but the announcement has drawn scrutiny because its published results do not establish that the system meets the company’s 2018 charter definition of AGI: a highly autonomous system that outperforms humans at most economically valuable work. Astra reportedly scored strongly on FrontierMath Tier 4, ARC-AGI-3, and ExploitBench, yet OpenAI did not release results for GDPval, its benchmark for assessing performance on economically valuable tasks across 44 occupations.
The launch also raises assurance questions. Astra’s system card reportedly identifies it as OpenAI’s most aligned model while showing substantially lower chain-of-thought monitorability than prior models, despite improved action-only monitoring. Separately, OpenAI and Microsoft reportedly amended their agreement so revenue-sharing continues through 2030 regardless of technological progress, removing AGI as a contractual trigger; critics contend that benchmark gains and launch messaging do not demonstrate reliable, verifiable performance on long-duration real-world work.

Track how attackers are adapting to this technology.
11 events from the most recent confirmed update back to the earliest known activity.
At the GPT-6 Astra briefing, OpenAI president Greg Brockman ended by saying, “Welcome to the AGI era,” and said he personally believed OpenAI had reached AGI. OpenAI reported Astra scores of 98% on FrontierMath Tier 4, 99.9% on ARC-AGI-3, and 100% on ExploitBench, while its launch materials did not include GDPval results.
OpenAI confirmed that it had made a confidential S-1 submission.
OpenAI and Microsoft amended their agreement so OpenAI's revenue-sharing payments reportedly continue through 2030 regardless of technological progress, removing AGI as a contractual trigger for ending those payments.
OpenAI closed a funding round at a reported post-money valuation of $852 billion.
OpenAI reportedly declared an internal code red after Google's Gemini 3 outperformed it on benchmarks.
OpenAI introduced GDPval, a benchmark intended to test economically valuable tasks across 44 occupations.
Sam Altman said that AGI was not a super useful term.
According to reporting cited in the source, OpenAI's 2023 Microsoft contract defined AGI using a threshold of at least $100 billion in profit.
OpenAI's 2018 charter defined AGI as highly autonomous systems that outperform humans at most economically valuable work.
OpenAI stated that unsafeguarded GPT-6 Astra testing discovered and exploited two previously unknown vulnerabilities and demonstrated exploit development against hardened browsers and operating systems. OpenAI said the production model restricts advanced offensive cybersecurity tasks while it plans broader defensive access through Daybreak.
A Medium article alleged that OpenAI classified Astra as “critical” for cybersecurity risk under its Preparedness Framework after preliminary evaluations of autonomous cyber capabilities. It reported that advanced cyber features were restricted to alpha testers and a purported Daybreak Blue defensive-security program, with refusal training, monitoring, and session-disruption controls.
Vulnerabilities, threat actors, malware, products, organizations, and breaches Mallory has linked to this story.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
4 references tracked. Mallory keeps watching after this page renders.
infoq.com
Open sourcemedium.com
Open sourcetechtrenches.dev
Open sourceopenai.com
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.