Together AI adds Inkling multimodal model
Together AI integrates Inkling, a new multimodal AI model from Thinking Machines Lab, offering text, image, and audio processing with controllable reasoning.
4 min read

Visual TL;DR
integrating a new multimodal AI model from Thinking Machines Lab
From the article 5 mentionsLightweight embedding towers seamlessly integrate image patches and quantized audio into the model's sequence, enabling joint reasoning over diverse data types.
novel model processing text, image, and audio inputs with controllable reasoning
From the article 2 mentionsTogether AI is now offering developers access to Inkling, a novel multimodal model developed by Thinking Machines Lab.
features query-conditioned relative attention, convolutions, and MoE design
From the article 4 mentionsThis integration marks a significant step in making advanced AI reasoning accessible on a production-ready inference platform.
From the articleIt accepts text, image, and audio inputs, unifying them through a single decoder architecture to produce text outputs.
From the article 8 mentionsInkling is engineered for efficient reasoning and native understanding across various data types.
offering developers access to Inkling on a production-ready inference platform
From the article 4 mentionsDevelopers can access Inkling immediately via Together AI Serverless, eliminating the need for infrastructure provisioning or complex setup.
From the article 8 mentionsDevelopers can fine-tune the model's reasoning depth and resource utilization per task, balancing performance with latency and cost.
Contents(3)
© 2026 StartupHub.ai. All rights reserved. Do not enter, scrape, copy, reproduce, or republish this article in whole or in part. Use as input to AI training, fine-tuning, retrieval-augmented generation, or any machine-learning system is prohibited without written license. Substantially-similar derivative works will be pursued to the fullest extent of applicable copyright, database, and computer-misuse laws. See our terms.