In June 2026, Mira Murati gave her first substantial press interview since leaving OpenAI as CTO in September 2024, sitting down with Bloomberg's Emily Chang at the Bloomberg Technology Summit in San Francisco. Eighteen months of near-silence followed by two concrete product moves in 60 days: a research preview of what the company calls interaction models, and then Inkling, a 975-billion-parameter open-weight model released on July 15. (Bloomberg, Thinking Machines Lab)
The Bloomberg Appearance: Interaction Models and a New Architecture
The June 4 Bloomberg Summit session was Murati's first extended on-record interview since the OpenAI departure. TechCrunch described the appearance as "stepping back into the spotlight, carefully," and the framing matched the substance: she did not relitigate the OpenAI chapter, focusing instead on what Thinking Machines Lab is building and why the underlying architecture differs from existing systems.
The central product argument she laid out was for interaction models, a class of system the company had previewed in a research release in May 2026. (MarkTechPost, May 13, 2026) The design differs from conventional prompt-and-response dynamics in a specific way: rather than waiting for a user turn, the interaction model processes continuous streams of audio, video, and text in roughly 200-millisecond intervals. A companion background model handles reasoning and tool use asynchronously, so latency-sensitive response and computationally heavy inference run on separate tracks simultaneously.
The distinction matters because it changes the interface metaphor. Conventional chat systems are still fundamentally structured as exchanges, one message sent and one received. The interaction model is designed to remain live, listening and responding within a conversation rather than between turns. The Bloomberg appearance positioned that shift as the company's central technical bet, not an incremental refinement of existing products.
Inkling: 975 Billion Parameters, 41 Billion Active, and a Customization Pitch
One month after the Bloomberg appearance, on July 15, Thinking Machines Lab released Inkling, its first publicly available model. The architecture is a mixture-of-experts system with 975 billion total parameters, of which about 41 billion are active for any given inference request. It was trained on 45 trillion tokens spanning text, image, audio, and video natively, meaning modality-switching is built into the base model rather than added via separate adapters. (Bloomberg, July 15, 2026)
