OpenAI's probability of releasing a new model in August has dropped from baseline to 13% — and this week, two facts point to a single judgment: AI frontier firms are no longer just talking about safety, they're putting real money on the brakes. OpenAI paused roughly two weeks of frontier reinforcement learning training, while Anthropic simultaneously restricted release of its most advanced model, Mythos. OpenAI safety lead Mia Glaese publicly framed it: a return to training remains "quite far off."
What this is
OpenAI's Preparedness Framework received a major upgrade this week: "critical cybersecurity capabilities" have been moved from a pre-deployment evaluation metric to a hard constraint during the development process. Critical cybersecurity capabilities, in plain terms, mean a model that can develop zero-day exploits (previously undisclosed security vulnerabilities) on hardened systems or execute end-to-end attacks without human intervention. After OpenAI's internally code-named Astra model crossed this threshold, a significant portion of Astra's workloads had to migrate to higher security standards before continuing. Anthropic moved in lockstep, prioritizing cybersecurity patches before considering public release.
Industry view
Supporters argue this is the responsible call — capability advancement and safety controls are different workloads on the same infrastructure, and post-deployment remediation is far more expensive. Anthropic flagged the possibility of "recursive self-improvement" (AI iteratively upgrading itself) back in June and advocated proactively slowing down.
But VC David Sacks publicly pushed back sharply: in practice, review mechanisms raise the industry bar. Only well-capitalized frontier firms can meet the scrutiny, compliance, and safety demands; open-source models scattered globally can't be uniformly regulated. The likely outcome is not a safer field — it's a handful of companies gaining larger advantage. Worth noting: Anthropic is now valued near $1 trillion and has filed IPO paperwork, so the "opportunity cost" of voluntarily slowing down is non-zero — competitors haven't stopped.
Impact on regular people
For enterprise IT: access windows to frontier models may slip. Projects expecting new capabilities in Q3 need to re-plan timelines.
For individual careers: AI tooling iteration pace is slowing in the short term. Workflow improvements that depend on new model rollouts can breathe a little.
For consumer markets: if "safety as moat" logic hardens among frontier firms, alternatives from smaller vendors and the open-source ecosystem may get further marginalized.