Thinking Machines Lab debuts Inkling, its first open model
The lab's inaugural open-weights release is a mixture-of-experts system that takes image and audio inputs, shipped under a permissive Apache 2.0 license.

Thinking Machines Lab has released Inkling, described as its first open-weights model and a mixture-of-experts (MoE) multimodal system that accepts image and audio inputs alongside text. The model is published under the permissive Apache 2.0 license, according to the lab's announcement.
The MoE design is notable because it activates only a subset of the model's parameters for any given input, a routing approach that lets teams scale total capacity while keeping the compute cost of each forward pass in check. Pairing that architecture with native image and audio handling puts Inkling in the growing category of open models built to reason across more than one input type.
Why it matters
An Apache 2.0 license is one of the least restrictive terms an AI lab can choose, which means developers can study, fine-tune, and deploy Inkling commercially without the usage caveats attached to many so-called open releases.
- Open weights: the model parameters are available for download and self-hosting.
- Multimodal inputs: image and audio are supported in addition to text.
- MoE architecture: sparse expert routing aimed at efficiency at scale.
As an initial release, Inkling establishes a baseline rather than iterating on a prior version. The full picture — including parameter counts, context length, and benchmark results — will come into focus as the community begins testing the weights against comparable open multimodal systems.
Sources
More in Any-to-Any
StepFun's StepAudio 3 Realtime targets live voice AI
The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.
SenseTime's SenseNova-U1.5 Unifies Vision Tasks
An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

SenseTime Releases SenseNova U1.5 8B Any-to-Any Model
The new 8B multimodal model handles text, images, and image editing within a single native architecture.
0 comments
No comments yet. Be the first to weigh in.