On July 15, 2026, Thinking Machines Lab published Inkling, a 975-billion-parameter open-weight model, releasing full weights under the Apache 2.0 licence and positioning it as a structural alternative to the proprietary APIs sold by OpenAI and Anthropic. The move crystallises a bet that former OpenAI CTO Mira Murati has been building toward since founding the company in early 2025: that enterprises will pay more for customisable infrastructure than for commodity intelligence delivered by the token.
The open-weight wager: 975 billion parameters, zero licensing fees
Inkling is a Mixture-of-Experts (MoE) model: 975 billion total parameters, of which 41 billion are active per token. The company trained it on 45 trillion tokens spanning text, images, audio, and video; the model processes all four modalities as input and returns text, with a 1-million-token context window. A smaller companion model, Inkling-Small, was previewed alongside the main release: 276 billion total parameters, 12 billion active per token.
