The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
New TodayDeepSeek

DeepSeek Releases V4-Pro, an MIT-Licensed MoE Model

The company's newest flagship targets reasoning and coding while keeping a permissive open-source license.

Aug 13, 2026
Text / LLMReasoningCode
Read the story25K downloads
DeepSeek-V4-Pro-0813
Size
1.7 TB
Precision
INT8
Architecture
DeepseekV4ForCausalLM

Muse Glimmer 30B
Meta AI/Vision-Language

Meta's Muse Glimmer 30B Targets Local Agentic Coding

A 30-billion-parameter multimodal model built to run locally, released under Apache 2.0 with an eye on agentic coding workflows.

Aug 10, 2026
Qwen3.8-2.4T-A95B
Qwen · Alibaba/Text / LLM

Qwen releases 2.4T-parameter open MoE with 95B active

Alibaba's Qwen team pushes its largest sparse model yet, activating 95 billion parameters per token from a 2.4-trillion-parameter pool.

Aug 8, 2026
Qwen3.8-27B
Qwen · Alibaba/Vision-Language

Qwen3.8-27B Brings Vision to a Dense Model

Alibaba's latest Qwen release pairs image understanding with text under a permissive Apache-2.0 license.

Aug 5, 2026

The Weekly Weights

Every open release, in your inbox.

One concise email a week: the model launches that mattered, leaderboard shifts, and what's coming next. No noise.

The feed

Latest releases

View all
XingChen AGI/Text / LLM

Xing4.0 arrives as a 29B MoE with 4B active params

inclusionAI's new text model uses a mixture-of-experts design to keep compute low while shipping under an Apache-2.0 license.

Sep 16, 2026
Text / LLM
Xing4.0-29B-A4B
StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
Text → SpeechSpeech → Text
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
Vision-LanguageReasoning
Agnes-3.0-Flash
Internlm/Reasoning

InternLM's Atria Dawn Preview Targets Agentic Tasks

A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.

Sep 11, 2026
ReasoningText / LLM
Atria Dawn Preview
StepFun/Music

StepFun's StepAudio 3 Music Plans Before It Plays

The new open model separates musical structure from sound, generating long-form tracks from text prompts with an explicit planning stage.

Sep 10, 2026
Music
StepFun/Text → Speech

StepFun's StepAudio 3 Gen Unifies TTS and Music

A single discrete autoregressive model handles speech, voice design, sound effects, and music generation.

Sep 10, 2026
Text → SpeechMusic

Momentum

Trending models

View all
1

Zhipu AI

GLM-5.3-Flash

Text / LLM · 2.4M downloads

+1.3M
2

DeepSeek

DeepSeek-V3.2

Text / LLM · 2.2M downloads

+508.8K
3

Google DeepMind

Gemma 4 E2B

Any-to-Any · 3.6M downloads

+414.1K
4

rednote-hilab

dots.ocr

Vision-Language · 710.9K downloads

+342.2K
5

DeepSeek

DeepSeek-V4-Flash-Vision-Exp

Vision-Language · 668.3K downloads

+224.4K

Benchmarks

Open LLM leaderboard

All boards
#ModelAvg.
1
MaziyarPanahi/calme-3.2-instruct-78b
52.1
2
MaziyarPanahi/calme-3.1-instruct-78b
51.3
3
dfurman/CalmeRys-78B-Orpo-v0.1
51.2
4
MaziyarPanahi/calme-2.4-rys-78b
50.8
5
huihui-ai/Qwen2.5-72B-Instruct-abliterated
48.1
6
Qwen/Qwen2.5-72B-Instruct
Qwen · Alibaba
48.0
By type

Newest in each category

View all
Text → Image →
SenseTime/Any-to-Any

SenseTime's SenseNova-U1.5 Unifies Vision Tasks

An 8B model drops the usual encoder and VAE in favor of a single native architecture spanning understanding, reasoning, and image generation.

Sep 9, 2026
Code →
Tencent/Reasoning

Tencent's T1 Targets Long-Horizon Terminal Work

A 122B mixture-of-experts model trained with reinforcement learning claims state-of-the-art results on Terminal-Bench.

Sep 9, 2026
Text → Video →
MiniMax-H3
MiniMax/Text → Video

MiniMax Releases H3 Video Model on Hugging Face

The company's new diffusion model handles text-to-video and image-to-video, with support for joint audio-video generation.

Jul 28, 2026
Vision-Language →
Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
On the horizon

Upcoming launches

View all
AnnouncedAug 26, 2026

Qwen Teases 3.8-Flash-Next, a 125B Sparse MoE

Alibaba's next Qwen release pairs a large parameter pool with a tiny active footprint, promising speed without the full compute bill.

Qwen · Alibaba·Text / LLM
AnnouncedSep 9, 2026

DeepSeek Announces V4.1 Flash, a Cheaper Reasoning Model

The new mixture-of-experts model is billed as more capable than V4 Pro while costing less to run, and ships under an MIT license.

DeepSeek·Text / LLM
Announced

Zhipu Confirms 'Ox Alpha' Is a New GLM Model

The Chinese AI lab says its stealth-tested system belongs to the GLM series and will ship with open weights under an MIT license.

Zhipu AI·Text / LLM