The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDots Studio3 preview
Dots StudioAny-to-Any

dots3-note preview brings audio and vision to one model

An early build of a multimodal, long-context agentic system arrives on Hugging Face with support for both images and sound.

Aug 9, 2026
UpdateOther
dots3-note (preview)

dots-studio has published an early preview of dots3-note, a multimodal model designed to handle images, audio, and text within a single long-context, agentic framework. The release is available now on Hugging Face.

The model is described as "multimodal-any," meaning it aims to accept a range of input types rather than specializing in a single one. Vision-language capabilities sit alongside audio understanding, positioning dots3-note as a general-purpose system for tasks that blend perception with reasoning.

What we know

  • Multimodal support spanning vision and audio, plus text
  • A long-context, agentic design intended for multi-step tasks
  • Distributed under a non-standard "other" license
  • Published as a preview, so specifications may still shift

Key technical details remain unpublished at this stage. The record lists no confirmed parameter count, context length, or benchmark results, and the license is marked simply as "other" — worth checking carefully before any commercial use.

Why it matters: Combining audio and vision in one openly available agentic model is still relatively uncommon, and preview releases like this offer early hands-on access for developers experimenting with cross-modal workflows. As with any preview, expect the details to firm up as the family matures.

Sources

  • dots-studio/dots3-note-prev

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size576.9 GB
PrecisionBF16
ArchitectureDots3NoteForCausalLM
LicenseOTHER
Downloads906
Likes285

Modalities

Any-to-AnyVision-Language

0 comments

No comments yet. Be the first to weigh in.

More in Any-to-Any

Clef
Cloudflare/Vision-Language

Cloudflare's Clef brings structured decisions to open models

The new open-weight vision-language family outputs typed, structured results and arrives alongside a reinforcement-learning fine-tuning platform.

Oct 1, 2026
FLUX 3 Action Base
Black Forest Labs/Any-to-Any

Black Forest Labs brings FLUX to robotics

The image-model maker's FLUX 3 Action Base extends its generative stack into world-action modeling, released through the LeRobot ecosystem.

Sep 22, 2026
MiMo-V2.6 (Flash/Pro
Xiaomi/Any-to-Any

Xiaomi expands MiMo line with V2.6 multimodal models

The new Flash, Pro, and Distill variants add vision, audio, agentic behavior, and long-context handling to Xiaomi's open MiMo family.

Sep 21, 2026