The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDots Studio3 preview
Dots StudioAny-to-Any

dots3-note preview brings audio and vision to one model

An early build of a multimodal, long-context agentic system arrives on Hugging Face with support for both images and sound.

Aug 9, 2026
UpdateOther
dots3-note (preview)

dots-studio has published an early preview of dots3-note, a multimodal model designed to handle images, audio, and text within a single long-context, agentic framework. The release is available now on Hugging Face.

The model is described as "multimodal-any," meaning it aims to accept a range of input types rather than specializing in a single one. Vision-language capabilities sit alongside audio understanding, positioning dots3-note as a general-purpose system for tasks that blend perception with reasoning.

What we know

  • Multimodal support spanning vision and audio, plus text
  • A long-context, agentic design intended for multi-step tasks
  • Distributed under a non-standard "other" license
  • Published as a preview, so specifications may still shift

Key technical details remain unpublished at this stage. The record lists no confirmed parameter count, context length, or benchmark results, and the license is marked simply as "other" — worth checking carefully before any commercial use.

Why it matters: Combining audio and vision in one openly available agentic model is still relatively uncommon, and preview releases like this offer early hands-on access for developers experimenting with cross-modal workflows. As with any preview, expect the details to firm up as the family matures.

Sources

  • dots-studio/dots3-note-prev

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size576.9 GB
PrecisionBF16
ArchitectureDots3NoteForCausalLM
LicenseOTHER
Downloads1.2K
Likes227

Modalities

Any-to-AnyVision-Language

0 comments

No comments yet. Be the first to weigh in.

More in Any-to-Any

Mistral AI/Vision-Language

Mistral debuts Shieldstral, a 3B safety model

The open-weights multimodal moderation model brings content safety checks to both text and images under an Apache 2.0 license.

Aug 4, 2026
Nemotron VoiceChat 11B
NVIDIA/Text → Speech

NVIDIA's Nemotron VoiceChat 11B Targets Spoken AI

An 11-billion-parameter voice conversation model built atop Nemotron Nano 9B v2 arrives on Hugging Face.

Jul 29, 2026
SenseNova U1.5 8B MoT Preview
SenseTime/Any-to-Any

SenseTime Debuts SenseNova U1.5 8B Multimodal Preview

An 8B any-to-any model that reads, generates, and edits images at up to 4K resolution, released as an early preview.

Jul 28, 2026