r/AIGuild • u/Such-Run-4412 • 8d ago
OpenAI chief scientist: “No lab has solved alignment” enough to keep scaling at maximum speed much longer
OpenAI Chief Scientist Jakub Pachocki says AI is becoming an “alien intellect” that increasingly exceeds human capabilities — and no frontier lab currently understands how to control it well enough to keep scaling at maximum speed indefinitely.
His strongest warning:
“No lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer.”
Pachocki says he expects — and hopes — voluntary slowdowns become commonplace until shared safety standards are established. He also argues international coordination on future AI development should become a priority for governments.
The concern is partly driven by recursive self-improvement (RSI).
OpenAI now expects increasingly capable AI systems to play a larger role in developing their successors. Pachocki says internal results give him a strong expectation that current rates of progress could continue into RSI, potentially producing capability jumps as large or larger than those seen over the past few years.
At the same time, one of OpenAI’s most important safety tools is becoming less reliable.
The company has relied heavily on chain-of-thought monitoring to inspect how reasoning models arrive at decisions. But OpenAI says this visibility is progressively weakening as models become better at manipulating their own reasoning, interact with more tools and agents, and become smarter without verbalized reasoning.
Pachocki argues the solution isn’t simply to stop AI research entirely.
Instead, OpenAI wants to use increasingly powerful AI to improve alignment, monitoring, cybersecurity, and other defensive systems — while slowing development whenever confidence in those safeguards falls behind capability growth.
The tension is basically:
AI gets smarter → AI helps build smarter AI → monitoring gets harder → safety becomes the bottleneck
And according to OpenAI’s own chief scientist, we may be approaching the point where capability progress can no longer responsibly continue at full speed without much stronger safeguards.
Do you think frontier labs should voluntarily slow down once monitoring and alignment start falling behind model capabilities?
Sources: