On June 3, 2026, Fei-Fei Li published a taxonomy that assigns every system calling itself a "world model" to one of three functions: renderer, simulator, or planner. The piece, posted to Li's personal Substack, is the most precise public statement the World Labs CEO has made about what separates a genuine world model from a video generator with better marketing.
From Video Generator to Physics Engine: the Renderer-Simulator Gap
The sharpest point in Li's taxonomy is the distinction between a renderer and a simulator. A renderer takes inputs, whether text, image, or video, and produces pixels: a visual representation of what a scene looks like. Most AI systems currently described as world models, including video generators such as Sora, Runway, and Kling, are renderers by this definition. Li's taxonomy is explicit: systems that stop at the renderer stage are not true world models.
