The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestThinkingmachinesSmall
ThinkingmachinesVision-Language

Thinking Machines Debuts Inkling Small, a Compact Multimodal MoE

The Apache-2.0 model brings mixture-of-experts efficiency to image, audio, and text tasks in a smaller footprint.

Jul 27, 2026
NotableApache 2.0
Inkling Small

Thinking Machines has released Inkling Small, a compact variant of its Inkling multimodal model family. Published on Hugging Face under an Apache-2.0 license, the model is built as a mixture-of-experts (MoE) system designed to process image, audio, and text inputs.

The "Small" designation signals a lighter-weight member of the lineup, aimed at teams that want multimodal capability without the compute demands of a full-scale flagship. MoE architectures activate only a subset of parameters per token, which typically lets developers run larger effective models at a fraction of the inference cost — a practical fit for a smaller, more deployable release.

Why it matters

Open multimodal models that reach beyond image-and-text into audio remain relatively rare, and a permissive Apache-2.0 license makes Inkling Small easy to adopt commercially. A few things stand out:

  • Broad input support across image, audio, and text in a single model
  • MoE efficiency, favoring lower active compute at inference time
  • Apache-2.0 licensing, removing most barriers to commercial use

Thinking Machines has not published detailed parameter counts, context length, or benchmark figures alongside the release, so the model's precise capabilities and how it stacks up against peers remain to be seen. For now, the availability of an openly licensed, any-modality MoE is a meaningful addition to the growing catalog of accessible multimodal systems. Interested developers can find weights and any accompanying documentation on the model's Hugging Face page.

Sources

  • thinkingmachines/Inkling-Small

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size531.9 GB
PrecisionBF16
ArchitectureInklingForConditionalGeneration
LicenseAPACHE-2.0
Downloads483.4K
Likes407

Modalities

Vision-LanguageText / LLMAny-to-Any

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
inclusionAI/Vision-Language

LLaDA-UI Brings Diffusion Decoding to GUI Agents

inclusionAI's 16.7B MoE vision-language model uses block-wise diffusion to drive graphical interface tasks.

Sep 8, 2026