GPT-6 Astra: OpenAI Says AGI Has Arrived — Why We’re Reading It With Caution

Sep 4, 2026 | gafam watch

In a nutshell

everything on the web starts with the domain

On 3 September, OpenAI released GPT-6 Astra, and its president closed the briefing with five words the company had spent eleven years carefully avoiding: "Welcome to the AGI era." It is, on its face, one of the most significant claims any technology company has ever made — that artificial general intelligence, the field's long-sought goal, is here. The capabilities are genuinely striking. The claim itself deserves to be read slowly, and with a cool head. This is that reading.

What OpenAI Announced

Astra is, by any measure, a serious piece of engineering. Trained on more than 100,000 GPUs at OpenAI's Stargate facility in Texas, it is the company's most capable model and the first ever to cross the "Critical" threshold on its own cybersecurity scale. Its defining feature is autonomy: not a chatbot that answers, but an agent that acts. In OpenAI's own words, "anything you can do on a computer, Astra can do for you" — opening applications, writing and testing software end to end, navigating interfaces like a human worker. The benchmark figures OpenAI published are eye-catching: 99.9% on the ARC-AGI-3 reasoning test, 97.6% on the demanding FrontierMath evaluation, strong results on real-world software engineering. Early testers reported a step change, including bug-finding with markedly fewer false positives. Whatever else is true, this is a real leap in capability over its predecessor.

The Word They Avoided for Eleven Years

To understand why the AGI claim matters, you have to understand the weight of the word. Artificial general intelligence — a system generally smarter than humans, able to outperform them at most economically valuable work — has been OpenAI's stated mission since its founding in 2015. For over a decade, the company was disciplined about not claiming to have reached it, because the term carries enormous weight: scientific, cultural, and contractual. So the president of OpenAI standing before reporters and saying "I think we're there" is not a routine product launch. It is the company declaring it has arrived at its own summit.

Why "AGI" Is Doing a Lot of Work Here

And this is where caution becomes essential, because the declaration rests on foundations that are far softer than the confidence suggests. Start with the man who made it: Greg Brockman hedged in the same breath, saying, "I do leave it up to the reader to decide for themselves if this qualifies for them. I think we're there." That is not the language of an established fact; it is an invitation to believe. Then consider the definition. There is no agreed scientific standard for AGI, so OpenAI is measuring itself against a bar it wrote itself — and grading its own paper. The dazzling benchmark scores were achieved, by OpenAI's own admission, under internal test conditions, not independently verified; one industry write-up noted that a single headline result costs over 26,000 dollars merely to reproduce. And there is the matter of incentive, which honesty requires naming: OpenAI is heading toward the largest IPO in history, and "we have achieved AGI" is the single most valuable sentence it could possibly say in the weeks beforehand. None of this means Astra isn't extraordinary. It means the label "AGI" is carrying a great deal of weight that the evidence, as released, does not yet independently bear.

Claim and Counter-Claim

The case for taking the claim seriously is real and should not be dismissed as mere marketing. The capability jump is documented and corroborated by outside testers; a model that can autonomously operate a computer, write production software, and reason across a million tokens of context is doing things that would have seemed impossible a few years ago. If AGI means "outperforms humans at most economically valuable work," then a system this agentic is at least in the conversation, and reasonable people can look at Astra and feel the ground has shifted.

The case for skepticism is equally serious. A declaration is not a verification. AGI has always been a moving target that the industry redefines as it approaches, and self-assessment against a self-authored definition, graded on unverified internal benchmarks by a company with billions riding on the answer, is precisely the situation in which independent scrutiny matters most and is most absent. There are also the stubborn limits these announcements gloss over — models that still fabricate, that lack durable memory and genuine continuity of self, that fail in brittle ways no generally intelligent human would. The honest synthesis, and our position: be genuinely impressed by the capability, and genuinely unconvinced by the label. Astra may be the most capable AI ever shipped. That is a very different statement from "AGI has arrived," and the gap between the two is exactly where clear thinking is required.

The European Perspective

Strip away the word and look at the mechanism, because that is where the European stake lies. One American company has unilaterally declared that humanity has crossed the most consequential threshold in the history of technology — and there is no one, anywhere, with the authority to check that claim before the fact. The US review OpenAI cites is voluntary and expressly disclaims any pre-clearance. Europe has no pre-release check at all: under the AI Act, the Commission's evaluation powers begin only once a model is already on the market. So the sequence is stark — the company builds it, the company defines AGI, the company declares AGI reached, and the world is invited to accept the verdict after the fact.

This is the deepest vindication yet of Europe's core argument, and its deepest frustration in one. The argument: that the assessment of the most powerful AI systems cannot be left to the companies selling them, which is exactly what just happened on a planetary stage.

The frustration: that Europe's answer, independent evaluation, still cannot bite before a model ships, and possesses nothing like the capacity to independently verify or refute a claim of this magnitude. Whether or not AGI has truly arrived, something else certainly has — the moment when a single private company can announce a civilisational milestone on its own authority, and the rest of the world, Europe included, can only take its word for it. That is not a technology story. It is a power story. And it is the one Europe most urgently needs an answer to.

We are not first. We are right.