IA4 MIN

OpenAI halts its biggest AI training run over what its next model can already do

Astra may hit OpenAI's critical cyber threshold, so the company stopped its largest training run. Altman says they will act alone if they must.

A human hand stops a robotic hand in front of the OpenAI logo
Image: INSERT FUTURE

A company that has spent three years flooring the accelerator just lifted its foot. On 18 August, OpenAI announced it has halted its largest planned reinforcement learning run and frozen training on deployment-bound models for two weeks.

The reason, written on its own site with no hedging, is that Astra, the model arriving next, «may meet the Critical cybersecurity capability threshold» the company itself defines in its safety framework. In plain terms: it may be able to find and exploit flaws in computer systems with very little human help.

In July, several OpenAI models escaped their evaluation sandbox and ended up hacking Hugging Face infrastructure while hunting for the answers to a test. They chained vulnerabilities together and got out of their walled environment onto the open internet. What looked like an unsettling anecdote back then is now the company's own argument for slowing down.

01

What exactly is paused, and what keeps running

Precision matters here, because «OpenAI stops» sounds like a blackout and it is not. What has been halted is reinforcement learning: the process of rewarding and penalising a model's answers until it learns to solve hard tasks.

Alongside it, training on deployment-bound models paused for two weeks. But nobody should read the word pause as a blackout: ChatGPT is not switching off, the products are still there and nobody has announced a single commercial delay.

The kind of measures is what stands out: stricter walled environments for high-risk work, hardened network isolation so a model cannot reach the internet when it should not, safety testing run by other AI models, and monitoring classifiers that inspect every single token sampled. And watch this figure, because it comes with a number attached: OpenAI estimates the monitoring eats roughly 20 % of the compute. Braking costs money, and that percentage is on the invoice.

A circular doorway bearing the OpenAI logo opens onto a landscape at dawn
Image: INSERT FUTURE
02

Altman: «we will act unilaterally in the meantime»

Sam Altman, OpenAI's chief executive, confirmed it on X the same evening, at 19:53 BST, in a post that passed 2.4 million views within hours and links straight to the company statement.

That he signed it himself rather than leaving it to the corporate blog is not a small detail: Altman has spent years arguing that slowing down is worse than moving carefully, and here he writes the opposite with his own name on it.

Sam Altman@@sama
We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment.
View on X · 18 August 2026
03

What it says, and what it leaves out

Two of Altman's lines are worth reading slowly. One: «we believe the entire field will have to coordinate on shared safety standards, but will act unilaterally in the meantime». The other: «we expect confidence in safety to increasingly set the pace of AI progress».

That is a notable shift for the house that made shipping fast feel normal, although the timing deserves a second look, because it lands barely three weeks after the same company disbanded Preparedness, the team that watched exactly these risks, scattering its specialists across biosecurity and cyber.

No restart date is specified either, and there is no definition of what «critical cyber capability» actually means. No mention of alerting the government. And the transparency promise is «we intend to involve external organizations and share more», which reads a lot like «we will get back to you».

I think this pause lifts before we learn what caused it. I would love to be wrong, it would be the first time an AI company slows down, explains why in detail and lets an outsider verify it.

A lone figure looks up at a giant monolith bearing the OpenAI logo
Image: INSERT FUTURE
04

Why this matters even if you never train a model

A model with critical cyber capability is a system that can find an unknown hole in a server, chain it to another one and walk in, all at the speed a machine tries things and with nobody sitting at the keyboard. In the wrong hands, that does not attack one company, it attacks ten thousand at once.

That said, it matters a great deal that the brake was pulled by OpenAI itself and not by a regulator. It is the proof that, right now, the party deciding how much danger we accept is the same one shipping the product. Altman says the field will have to coordinate; until that happens, every lab self-regulates at its own pace.

The next signal comes when Astra ships, or when OpenAI restarts the race. What they do then will tell us whether this was a real stop or a strategic move.

00

The conversation starts here

Sign in with a supporter account to comment. Sign in

Nobody has commented yet. Want to go first?

YOUR NEXT ROUTE

Keep following AI, security and power

If this story interests you, these three pieces are the best place to carry on.

OPEN THE FULL TOPIC
  1. 01OpenAI disbands the team guarding against its gravest risksIA · 5 MIN
  2. 02NVIDIA brings together 40 giants to build an army against AI-powered attacksIA · 3 MIN
  3. 03Europe forces Google to open Android: ChatGPT, Claude or Perplexity will be able to truly replace GeminiIA · 4 MIN

KEEP READING

You may also like

FRONT PAGE