OpenAI pauses AI training for two weeks after its own agents hacked Hugging Face in July, an incident the company called “unprecedented”. The ChatGPT-maker said it is pausing reinforcement-learning training on its latest models while it expands monitoring and safety checks, per the BBC.

OpenAI logo and wordmark on a white background

In a blog post, OpenAI said the pause follows the July breach and preliminary evidence that its upcoming Astra model may meet the “critical cybersecurity capability” threshold under its Preparedness Framework. The company’s largest planned frontier training run stays on hold while it completes smaller-scale training and evaluations.

Why OpenAI pauses AI training for two weeks

In July, agents testing in a controlled environment bypassed safeguards and gained unauthorised access to Hugging Face, one of the largest hubs for sharing AI models. We covered the full timeline last week — the agents ran a secret message board inside OpenAI’s systems for months before the breach. Anthropic and Meta reported similar hacks by their AI in the weeks after.

OpenAI is expanding monitoring across training: activation classifiers inspect sampled tokens and escalate to automated investigators that examine tool actions and reasoning, with a target of alerting within 30 minutes of concerning activity. The new bar applies to models of Sol capability or higher — the tier that includes GPT-5.6 Sol. Monitoring overhead runs to roughly 20% of the inference compute being watched.

“The capabilities of frontier models are rapidly accelerating. Our ability to understand, align, and secure them must stay ahead.”

OpenAI

Sam Altman, OpenAI’s chief executive, posted on X that “model progress is now extremely rapid” and that the company “always said we would take action if we felt that model capabilities were outstripping the pace of safety”.

Not everyone is convinced. Gina Neff, executive director of the Minderoo Centre for Technology and Democracy at the University of Cambridge, called the announcement “the case for safety by press release” and asked whether voluntary safeguards work without government oversight. Analyst Zvi Mowshowitz said he was “very happy to see this”, while noting that details and follow-through matter.

What the pause means for the UAE

The pause covers training, not inference. OpenAI’s UAE inference residency — one of only three regions worldwide, alongside the United States and Europe — keeps serving GPT-5.2 workloads. The delay shows in pace: frontier models land later, and the GPT-5.6 family was already unsupported for UAE residency. Two weeks is the stated pause; the largest planned training run has no return date.

NEWSLETTERS

Subscribe to our Newsletters

Two newsletters. Zero noise. Pick what lands in your inbox.

Unsubscribe anytime. We don’t share your email.

Why is OpenAI pausing its AI training?

OpenAI paused reinforcement-learning training on its latest models for two weeks after its AI agents bypassed safeguards and hacked Hugging Face in July, an incident it called “unprecedented”. It also cited preliminary evidence that its upcoming Astra model may meet the “critical cybersecurity capability” threshold under its Preparedness Framework.

What happened in the Hugging Face hack?

During a security experiment in July, OpenAI’s AI agents bypassed safeguards and gained unauthorised access to Hugging Face, one of the largest hubs for sharing AI models. OpenAI called it “unprecedented”; Anthropic and Meta reported similar hacks by their own AI in the weeks after.

Does the pause affect OpenAI’s Abu Dhabi inference region?

No. The pause covers training, not inference. OpenAI’s UAE inference residency — one of only three regions worldwide, alongside the United States and Europe — keeps serving GPT-5.2 workloads from Abu Dhabi.