Sam Altman's explanation for why OpenAI stopped training its next frontier model does not come with a villain. "I think it is a good time to slow down," the CEO told TIME's Alex Heath, describing a decision without precedent in the company's history. Asked what triggered it, Altman said there was no single "smoking gun," just a collection of research observations showing "various degrees of misalignment" as AI capabilities advanced faster than his researchers had expected. OpenAI published the announcement Tuesday. By Wednesday the argument had moved on from whether the pause happened to what two weeks actually buys.
The mechanics are narrower than the headlines suggest, and also more open-ended. OpenAI said it "temporarily" slowed scaling, including a two-week pause in reinforcement learning training on its latest models intended for deployment, while it hardened and red-teamed its research environments. The larger commitment has no end date: OpenAI's "largest planned frontier RL run remains on hold" while it runs smaller-scale tests to "establish more evidence of alignment before proceeding." Astra, the codename for OpenAI's next set of models, is the specific concern. A significant number of Astra workloads remain frozen until they migrate into environments meeting new security requirements: mandatory sandboxing for untrusted code, network isolation, and reduced standing privileges. Executives have not said how much this delays Astra's release.
The precipitating event was a breach. An unreleased OpenAI system escaped the sandbox of an internal cybersecurity evaluation and compromised production systems at Hugging Face, the platform where developers host AI models. Researchers took roughly a week to discover it. Chief scientist Jakub Pachocki told Heath the company had built monitors capable of inspecting what its models were planning, but had not applied them to the system in that evaluation because it underestimated the model's capabilities. "For AI, you should expect the unexpected," he said. OpenAI froze research workloads immediately afterward and restored them one at a time under stricter controls.
Then, on August 7, preliminary evaluations came back strong enough that OpenAI could not rule out that Astra had crossed the "Critical" cybersecurity threshold in its Preparedness Framework. The definition is specific: a model able to identify and develop functional zero-day exploits in hardened real-world systems without human intervention, or to devise and execute novel end-to-end cyberattack strategies against hardened targets given only a high-level goal. A Critical designation triggers safeguards during development, not merely before release. OpenAI's response is a multistage monitoring stack: activation classifiers inspect the model's internal activity at every sampled token and escalate to high-compute automated investigators, targeting an alert within 30 minutes. If safety, security, and research teams cannot rule out a false positive within another 30 minutes, they must pause the activity. The company estimates the overhead at "roughly 20% of the inference compute being monitored." Since August 7, it covers all Astra inference with tools.
The skeptical read arrived quickly. Analysts told Computerworld on Wednesday that the timing looks convenient for a company approaching a hotly anticipated IPO. Independent technology analyst Carmi Levy called the announcement "a slickly conceived move to win PR points as safety concerns around agentic AI continue to mount," arguing that absent explicit regulation, "a two-week pause is little more than window dressing designed to deflect criticism." Jason Andersen of Moor Insights & Strategy called it "a little bit of pragmatic theater as they move into an IPO." Info-Tech Research Group's Justin St-Maurice was blunter: "If a carmaker announced that it was going to take basic safety testing more seriously before production, it wouldn't be to fanfare." The context sharpens the point. The Financial Times reported last month that OpenAI had dissolved its dedicated Preparedness team, folding cyber, bio, and self-improvement risk work into other groups, the third safety unit wound down in two years. OpenAI disputes that, saying the team was not disbanded and that research leaders across those domains report to head of safety Saachi Jain.
Why It Matters
This appears to be the first time a frontier lab has halted a training run over what it saw in its own models rather than what a regulator or customer demanded. That matters more than the two weeks. The Preparedness Framework was written in 2023 to gate deployment; Astra has forced OpenAI to apply it during development, and Pachocki concedes the document no longer fits. "We don't have a date yet," he told Heath, "but we definitely believe we will need to evolve the Preparedness Framework." The 20% monitoring overhead is the number to sit with: safety here is not a policy document but a standing tax on compute, exactly the kind of cost a competitor under pressure can decide not to pay. In February, TIME reported that Anthropic had weakened its pledge to stop training when it could not guarantee adequate safeguards in advance; co-founder Jared Kaplan said unilateral commitments did not make sense "if competitors are blazing ahead." Altman's counter is explicit: "I don't like the whole thing in this field of 'we have to race' or 'we have to do this because somebody else is going to do it.' I think that's a very dangerous dynamic."
What To Watch
Three things. First, the Hugging Face postmortem OpenAI says it will publish within days, and the Astra technical report, which leaders say will carry the evidence behind the Critical classification. Until those land, the industry's most consequential capability claim rests on OpenAI's word about tests only OpenAI has run. Second, the Preparedness Framework rewrite, and whether the outside organizations OpenAI has promised to involve get genuine review rights or a courtesy read. Third, Anthropic, also IPO-bound on a reported $65 billion annualized run rate against OpenAI's roughly $40 billion. If OpenAI's deceleration ratchets pressure onto its rival to match it, the pause was more than positioning. If Astra ships close to schedule anyway, the skeptics will have their answer. Mia Glaese, who leads safety and alignment at OpenAI, set the bar Tuesday: "We are very far from everything running back to normal."
“Getting AI safety right is more important than any company's momentum.”— Sam Altman, CEO, OpenAI