Morning Edition · Tuesday, September 8, 2026Published at 2:18 AM EDT · New York
Jakub Pachocki argues that no lab's safeguards can support full speed scaling for much longer, and he wants voluntary commitments turned into mandatory thresholds enforced by outside auditors or governments.

Jakub Pachocki, the chief scientist of OpenAI, published an essay titled "An Alien Mind" on 6 September, arguing that the industry should be prepared to decelerate. He wrote that competing to build the most capable systems regardless of the risk is indefensible once the scale of that risk is understood, and that labs should slow down together when their safety work cannot match the pace of their capability gains. Bloomberg reported that he urged "extreme caution" on the pace of development, and Decrypt summarized his position that no lab's current safeguards are adequate to keep building more powerful systems at full speed for much longer.
The concrete request is institutional rather than technical. Pachocki proposes that today's voluntary company commitments become mandatory safety thresholds, enforced by independent auditors, governments or international bodies. His stated worry is recursive self-improvement: models that run research projects and eventually improve themselves without a human involved in the process.
Two things are worth separating. OpenAI has announced no pause, no delayed launch and no change to its compute plans, so this is an argument made public by an executive, not a corporate decision. And a leading lab that proposes the rule usually finds it easier to meet than a smaller challenger does, which is the standard objection to safety-threshold proposals like this one. What is verified is that this kind of restriction is already happening in practice: OpenAI limited the strongest cyber capabilities of its Astra model to vetted testers, and Anthropic ships Mythos 5.1 only to organizations cleared through its Cyber Verification and Life Sciences Verification programs, currently inside the United States only.
Part of a tracked trend
Oversight and Evaluation Lag Accelerating AI Capabilities
Over the next 3-6 months, evidence mounts that governance, evaluation, and agent-safety methods are failing to keep pace with capability growth, driving investment in interpretability, agent-manipulation benchmarks, and institutional-reform proposals.
Start a discussion in Townsquare.
More from this edition
Incumbent labs that already publish preparedness frameworks and can absorb third-party audits, since mandatory safety bars enforced by auditors or governments raise the fixed cost of frontier work and protect the capital already committed to it.
The essay itself is verified and quoted accurately, but it is one executive's argument rather than a company decision, and reporting that OpenAI's own compute disclosures show continued expansion means "slowdown" describes a proposed rule set, not an announced change to any training run.
An open-source-intelligence read of how likely this story is true with its real nuance, not a judgment of any outlet. It assesses the claim, weighing independent and adversarial reporting. How we label confidence.
What this means
The mechanism to watch is access restriction, not computing power. When frontier capability is released through vetted-organization programs and regional limits rather than an open application programming interface, the developers who lose access are small companies and firms outside the United States, while the winners are large enterprises with compliance staff who can pass a verification review. That reshapes who can build on frontier tools well before any law arrives, and it pushes everyone left out toward open-weight alternatives.
What to watch
Observations to monitor, not financial advice.
Synthesized from: Polylog editors · Anthropic News
Comments
0No comments yet.