Introspection co-founder Roland Gavrilescu told AI Engineer the next frontier is not a better model or a better harness but the loop that keeps improving itself.
He and co-founder Julian Bright left xAI a few months ago after working on agent infrastructure and cloud agents there to build a standalone company around always-on, long-horizon tasks. The talk frames a 2026 blueprint: trust the loop, distill what it learns, and measure valued work per watt.
The punchline came early.
Gavrilescu traced the arc from RL-trained reasoning to harness engineering to loops that run without touching code, then pointed to the first viral example that was not a coding agent at all. An engineer named AJ used what was then called Clawbot, now OpenClaw, to scrape inventories, pull Reddit prices, ping dealers and play PDF quotes against each other until the price was verifiable and the car could be locked in. That pattern, what he called the loop is the product, was already selling out Mac minis when developers started buying dedicated boxes to keep agents always on.
That history has a longer tail. Gavrilescu rooted the idea in OODA loops, the observe-orient-decide-act cycle coined for US Air Force fighter pilots in the 1970s, and argued models have been trained with that cadence in mind: call tools, take observations, act again. Put strong signals and verifiable work at the end and you get workers like Claude Code agents. The quality of the signal and the verifier decides whether the loop actually succeeds. Feed the artifacts of one loop back as the signal for the next and you get continuous improvement.
