Beyond RLHF: The Future of AI Automation

Former OpenAI researcher Diogo Almeida argues that the current RLHF-based AI era is limited to assistance, and true automation requires a new approach.

8 min read
Diogo Almeida speaking at a podium with a slide titled 'What's next after RLHF?'
Diogo Almeida, speaker at AI Engineer World's Fair.· AI Engineer

Visual TL;DR. RLHF Dilemma due to Human Preference Limits. Human Preference Limits leads to Diogo Almeida's Argument. Diogo Almeida's Argument advocates Future: True Automation. RLHF Dilemma defines Current AI: Assistance. Current AI: Assistance contrasts with Future: True Automation. Future: True Automation requires Smarter Software. Future: True Automation is goal of TypeSafe AI's Vision.

  1. RLHF Dilemma: current AI excels at assistance, struggles with true automation tasks
  2. Human Preference Limits: RLHF optimizes for human feedback, not autonomous task completion
  3. Diogo Almeida's Argument: former OpenAI researcher advocates moving beyond RLHF for automation
  4. Current AI: Assistance: RLHF-based models like ChatGPT are good conversational assistants
  5. Future: True Automation: requires a new approach beyond human preference optimization
  6. Smarter Software: AI needs to perform complex tasks without constant human oversight
  7. TypeSafe AI's Vision: Almeida's company aims to build AI for genuine autonomous automation
Visual TL;DR
Visual TL;DR, startuphub.ai Diogo Almeida's Argument advocates Future: True Automation. Future: True Automation is goal of TypeSafe AI's Vision advocates is goal of RLHF Dilemma Diogo Almeida's Argument Future: True Automation TypeSafe AI's Vision From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Diogo Almeida's Argument advocates Future: True Automation. Future: True Automation is goal of TypeSafe AI's Vision advocates is goal of RLHF Dilemma Diogo Almeida'sArgument Future: TrueAutomation TypeSafe AI'sVision From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Diogo Almeida's Argument advocates Future: True Automation. Future: True Automation is goal of TypeSafe AI's Vision advocates is goal of RLHF Dilemma current AI excels at assistance, struggleswith true automation tasks Diogo Almeida's Argument former OpenAI researcher advocates movingbeyond RLHF for automation Future: True Automation requires a new approach beyond humanpreference optimization TypeSafe AI's Vision Almeida's company aims to build AI forgenuine autonomous automation From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai Diogo Almeida's Argument advocates Future: True Automation. Future: True Automation is goal of TypeSafe AI's Vision advocates is goal of RLHF Dilemma current AI excelsat assistance,struggles with true… Diogo Almeida'sArgument former OpenAIresearcheradvocates moving… Future: TrueAutomation requires a newapproach beyondhuman preference… TypeSafe AI'sVision Almeida's companyaims to build AIfor genuine… From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai RLHF Dilemma due to Human Preference Limits. Human Preference Limits leads to Diogo Almeida's Argument. Diogo Almeida's Argument advocates Future: True Automation. RLHF Dilemma defines Current AI: Assistance. Current AI: Assistance contrasts with Future: True Automation. Future: True Automation requires Smarter Software. Future: True Automation is goal of TypeSafe AI's Vision due to leads to advocates defines contrasts with requires is goal of RLHF Dilemma current AI excels at assistance, struggleswith true automation tasks Human Preference Limits RLHF optimizes for human feedback, notautonomous task completion Diogo Almeida's Argument former OpenAI researcher advocates movingbeyond RLHF for automation Current AI: Assistance RLHF-based models like ChatGPT are goodconversational assistants Future: True Automation requires a new approach beyond humanpreference optimization Smarter Software AI needs to perform complex tasks withoutconstant human oversight TypeSafe AI's Vision Almeida's company aims to build AI forgenuine autonomous automation From startuphub.ai · The publishers behind this format
Visual TL;DR, startuphub.ai RLHF Dilemma due to Human Preference Limits. Human Preference Limits leads to Diogo Almeida's Argument. Diogo Almeida's Argument advocates Future: True Automation. RLHF Dilemma defines Current AI: Assistance. Current AI: Assistance contrasts with Future: True Automation. Future: True Automation requires Smarter Software. Future: True Automation is goal of TypeSafe AI's Vision due to leads to advocates defines contrasts with requires is goal of RLHF Dilemma current AI excelsat assistance,struggles with true… Human PreferenceLimits RLHF optimizes forhuman feedback, notautonomous task… Diogo Almeida'sArgument former OpenAIresearcheradvocates moving… Current AI:Assistance RLHF-based modelslike ChatGPT aregood conversational… Future: TrueAutomation requires a newapproach beyondhuman preference… Smarter Software AI needs to performcomplex taskswithout constant… TypeSafe AI'sVision Almeida's companyaims to build AIfor genuine… From startuphub.ai · The publishers behind this format

Diogo Almeida, a former key figure at OpenAI and co-author of influential papers on GPT-4 and ChatGPT, presented a compelling argument at the AI Engineer World's Fair, suggesting that the AI industry is poised to move beyond the era defined by RLHF. Almeida, now leading TypeSafe AI, posited that while RLHF has been instrumental in creating conversational AI assistants like ChatGPT, its inherent design for human preference optimization limits its utility for true automation.

Beyond RLHF: The Future of AI Automation - AI Engineer
Beyond RLHF: The Future of AI Automation — from AI Engineer

The RLHF Dilemma: Assistance vs. Automation

Almeida opened his talk by acknowledging the rapid advancements in AI, noting how benchmarks are consistently surpassed. However, he highlighted a dichotomy in current AI applications: tasks that are 'too good to be true' (like instruction following and chatbots) versus those that are 'too bad to be useful' (like customer service requiring human oversight or data entry). He argued that the simplest explanation for this divide lies in the core design of RLHF.

"Today's AI, everything inherited from RLHF, is incredible at the human in the loop stuff, but not for automation tasks," Almeida stated. He elaborated that RLHF's primary goal is to please the human user, which is ideal for assistance but not for autonomous tasks where error-free execution is paramount. The overpromising nature of current AI, he contended, is a feature of RLHF, designed to optimize for engagement rather than calibrated performance.

The Limitations of Human Preference Optimization

Almeida explained that RLHF works by collecting human preferences and then optimizing models based on those preferences. This process, while effective for creating helpful and engaging AI assistants, inherently leads to a gap between human preference and actual task results. He illustrated this with an anecdote about sending ChatGPT fart sound effects and asking for a musical critique, where the AI provided an elaborate, albeit nonsensical, response, demonstrating its bias towards generating a human-pleasing output.

He emphasized that the goal for automation is not to mimic human preferences but to execute tasks correctly and reliably, ideally operating in the background without human intervention. This fundamental difference, he argued, explains why current AI is adept at conversational tasks but falters in domains requiring precision and autonomy.

The Future: True Automation and Smarter Software

Almeida’s core thesis is that the next frontier for AI is true automation, which will require a new approach beyond RLHF. He suggested that the current SaaS landscape, despite AI advancements, has seen little fundamental change, with chatbots often being a superficial addition. He cited Garry Tan's phrase, "We're entering the golden age of just-in-time software," but cautioned that this could be a double-edged sword, leading to more software generation rather than inherently smarter software.

"What I want is smarter software," Almeida declared. "Why can't like B2B SaaS just be more expressive? Like, why are the like, the building blocks of software actually still the same?" He proposed that the AI industry should focus on redesigning the AI stack for reliability and automation, moving towards a future where AI can handle complex, rote tasks autonomously.

TypeSafe AI's Vision

Almeida revealed that his current venture, TypeSafe AI, is working on this very problem. Their core question revolves around what would happen if the AI stack were redesigned for reliability and automation. He expressed excitement about this work, stating, "I think it's one of the most satisfying things I've worked on." He also hinted at an upcoming announcement, noting that "the original scaling laws were incorrect."

The talk concluded with a call to action for those interested in building smarter software to sign up for TypeSafe's mailing list or careers page, and to follow him on Twitter for further insights.

© 2026 StartupHub.ai. All rights reserved. Do not enter, scrape, copy, reproduce, or republish this article in whole or in part. Use as input to AI training, fine-tuning, retrieval-augmented generation, or any machine-learning system is prohibited without written license. Substantially-similar derivative works will be pursued to the fullest extent of applicable copyright, database, and computer-misuse laws. See our terms.