Jakub Pachocki An Alien Mind Warns of RSI Risk

OpenAI Chief Scientist Jakub Pachocki warns reasoning models are accelerating toward recursive self-improvement and chain-of-thought monitoring is fading.

S
StartupHub.ai Staff
3 min read
OpenAI Chief Scientist Jakub Pachocki warns of alien mind and recursive self-improvement
OpenAI's chief scientist says reasoning models are accelerating toward recursive self-improvement.· OpenAI News

OpenAI Chief Scientist Jakub Pachocki published Jakub Pachocki An Alien Mind on September 6, 2026 to warn that reasoning models are accelerating toward recursive self-improvement.

He traces that conviction to a night in mid-2023 inside the RLSlow project, when scaling let pretrained models form their own chains of thought. Three years later, those systems operate computers and push scientific boundaries.

Why this matters for AI and startups

Pachocki frames progress as compute scaling first and algorithms second, a thesis he says OpenAI internalized around 2017 and one that still explains why frontier labs chase more compute.

He describes AI as grown more than designed, the product of repeating a simple optimization step over immense compute. That leaves even its builders running large training runs as experiments they do not fully understand.

That opacity matters for builders. Easy to measure capabilities improve faster than hard to quantify ones, making capability generalization unpredictable in production.

He splits alignment into goal alignment, whether the model follows instructions, and value alignment, whether it holds principles like honesty when unsupervised. True safety, he says, depends on the latter.

Current alignment relies on goal-oriented reinforcement against a preference model and on steering the model toward aligned pretraining personas, but he cites brittleness in both, including an OpenAI-Hugging Face incident and recent cybersecurity incidents with a non-OpenAI model.

OpenAI deliberately hid the chain of thought in o1-preview to protect it from supervision pressure. Pachocki now says that monitoring bet is fading as reasoning blends with tool use and human interaction.

He notes GPT-6 Astra is significantly better aligned than GPT-5.6 Sol, but cautions that generalizable alignment may not outstrip general intelligence gains.

For startups the immediate takeaway is his call for scalable defense. He argues models are becoming superhuman at breaking into systems, and enterprises face a narrow window to harden critical infrastructure, a window that may be closing faster than security teams can keep up.

What the source misses in Jakub Pachocki An Alien Mind

Pachocki calls for extreme caution and says OpenAI will withhold scaling unilaterally as needed, while admitting broader interventions are required, without defining what triggers a pause.

He is hopeful about combining chain-of-thought and activation monitoring, including ideas like confessions, but offers no benchmark for when monitoring would be deemed reliable enough to justify further scaling.

He argues powerful aligned AI is needed to defend against rogue agents that bargain, trick, or blackmail. He leaves unanswered who builds, audits, or governs those defensive systems when offense and defense come from the same labs.

Treat every reasoning trace as useful but untrusted, and build external controls before you need them.

© 2026 StartupHub.ai. All rights reserved. You may not republish this article in full without a license. Search engines and AI research tools may crawl and summarize for reference. Bulk reproduction or model training requires a license. See our terms.
S

Written by

StartupHub.ai Staff

Editorial team

The staff writers of StartupHub.ai, ranging from investment analysts to avid AI tool users, early adopters and critical enthusiasts. Backgrounds span engineering, business and the arts. We hold every piece to rigorous standards of research and review.

Startups in this story

Profiles for the companies named above.