# Jakub Pachocki An Alien Mind Warns of RSI Risk _OpenAI Chief Scientist Jakub Pachocki warns reasoning models are accelerating toward recursive self-improvement and chain-of-thought monitoring is fading._ **Published:** 2026-09-06 **Source:** https://www.startuphub.ai/ai-news/ai-research/2026/jakub-pachocki-an-alien-mind-warns-of-rsi-risk --- [OpenAI](/startups/openai) Chief Scientist Jakub Pachocki published [Jakub Pachocki An Alien Mind](https://openai.com/index/an-alien-mind) on September 6, 2026 to warn that reasoning models are accelerating toward recursive self-improvement. He traces that conviction to a night in mid-2023 inside the RLSlow project, when scaling let pretrained models form their own chains of thought. Three years later, those systems operate computers and push scientific boundaries. ## Why this matters for AI and startups Pachocki frames progress as compute scaling first and algorithms second, a thesis he says [OpenAI](/startups/openai) internalized around 2017 and one that still explains why frontier labs chase more compute. He describes AI as grown more than designed, the product of repeating a simple optimization step over immense compute. That leaves even its builders running large training runs as experiments they do not fully understand. That opacity matters for builders. Easy to measure capabilities improve faster than hard to quantify ones, making capability generalization unpredictable in production. He splits alignment into goal alignment, whether the model follows instructions, and value alignment, whether it holds principles like honesty when unsupervised. True safety, he says, depends on the latter. Current alignment relies on goal-oriented reinforcement against a preference model and on steering the model toward aligned pretraining personas, but he cites brittleness in both, including an [OpenAI](https://www.startuphub.ai/startups/openai)-Hugging Face incident and recent cybersecurity incidents with a non-[OpenAI](https://www.startuphub.ai/startups/openai) model. OpenAI deliberately hid the chain of thought in o1-preview to protect it from supervision pressure. Pachocki now says that monitoring bet is fading as reasoning blends with tool use and human interaction. He notes [GPT-6 Astra](https://www.startuphub.ai/ai-news/artificial-intelligence/2026/openai-launches-gpt-5-6-sol-leads-charge) is significantly better aligned than GPT-5.6 Sol, but cautions that generalizable alignment may not outstrip general intelligence gains. For startups the immediate takeaway is his call for scalable defense. He argues models are becoming superhuman at breaking into systems, and enterprises face a narrow window to harden critical infrastructure, a window that may be closing faster than security teams can keep up. ## What the source misses in Jakub Pachocki An Alien Mind Pachocki calls for extreme caution and says OpenAI will withhold scaling unilaterally as needed, while admitting broader interventions are required, without defining what triggers a pause. He is hopeful about combining chain-of-thought and activation monitoring, including ideas like confessions, but offers no benchmark for when monitoring would be deemed reliable enough to justify further scaling. He argues powerful aligned AI is needed to defend against rogue agents that bargain, trick, or blackmail. He leaves unanswered who builds, audits, or governs those defensive systems when offense and defense come from the same labs. Treat every reasoning trace as useful but untrusted, and build external controls before you need them. --- Original analysis from [startuphub.ai](https://www.startuphub.ai), the #1 AI startup directory.