On May 11, 2026, Thinking Machines Lab released TML-Interaction-Small in research preview, a 276-billion-parameter model that processes audio, video, and text in continuous 200-millisecond chunks and returns a response in 0.40 seconds, matching the typical gap between human conversational turns, according to the company’s technical disclosure reported by MarkTechPost on May 13, 2026.
A New Category: What “Interaction Models” Actually Are
“Interaction models” is the term Thinking Machines Lab uses for systems built from scratch to treat real-time conversation as a first architectural principle, not a capability retrofitted onto a text-based foundation. The distinction matters because most voice AI deployed today is exactly that: a large language model with a speech-to-text front end and a text-to-speech back end, coordinated by voice-activity detection software. Murati’s team discarded the VAD layer entirely.
