The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.

Category · text

Latest Text / LLM models

Open-weight large language models for chat, writing, and general reasoning — the foundation models you can download, fine-tune, and self-host instead of calling a closed API.

Filter

153 releases

Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
Text / LLMReasoning
Agnes-3.0-Flash
Internlm/Reasoning

InternLM's Atria Dawn Preview Targets Agentic Tasks

A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.

Sep 11, 2026
Text / LLMReasoning
Atria Dawn Preview
Unknown/Reasoning

ZGCM-1 arrives as a fully open 7B reasoning model

A compact foundation model targets math reasoning and agentic search with tool use, and its makers are releasing it fully open.

Sep 10, 2026
Text / LLMReasoning
Yandex/Text / LLM

Yandex releases a sparse T5 model with 0.6B active params

The AliceAI-T5-35B-A0.6B is an encoder-decoder mixture-of-experts model that keeps only a sliver of its 35B parameters active per token.

Sep 10, 2026
Text / LLM
AliceAI-T5-35B-A0.6B
Tencent/Reasoning

Tencent's T1 Targets Long-Horizon Terminal Work

A 122B mixture-of-experts model trained with reinforcement learning claims state-of-the-art results on Terminal-Bench.

Sep 9, 2026
CodeText / LLM
Edge0/Text / LLM

Edge0's 35B MoE Aims for SSD-Backed Edge Inference

A preview mixture-of-experts model uses trained routing prediction to run on machines that can't hold it all in memory.

Sep 8, 2026
Text / LLM
Edge0-35B-A3B-preview
Nex Agi/Vision-Language

Nex-N2.5-Pro arrives as an Apache-2.0 MoE vision model

A permissively licensed multimodal mixture-of-experts model built on a Qwen3-style MoE backbone.

Sep 8, 2026
Text / LLMVision-Language
Nex-N2.5-Pro
Nex Agi/Text / LLM

Nex AGI releases Nex-N2.5-mini, an open MoE multimodal model

The compact mixture-of-experts model handles both text and vision, and ships under a permissive Apache-2.0 license.

Sep 8, 2026
Text / LLMVision-Language
Nex-N2.5-mini
Thesysdev/Text / LLM

OUI-1: A Gemma-based diffusion model for generative UI

Thesys releases an experimental diffusion language model aimed at turning prompts into user interfaces, built atop Google's Gemma.

Sep 7, 2026
CodeText / LLM
OUI-1
OpenBMB/Text / LLM

OpenBMB's MiniCPM5-2B targets on-device AI

The compact 2-billion-parameter model adds long-context handling and tool-calling in a footprint small enough to run locally.

Sep 6, 2026
Text / LLM
MiniCPM5-2B
TokenRhythm/Text / LLM

NeoHorse-1-4B tunes Qwen3.5 for agentic work

A compact 4-billion-parameter model built for tool use, coding, and multi-step reasoning arrives from TokenRhythm.

Sep 5, 2026
CodeText / LLM
NeoHorse-1-4B
inclusionAI/Vision-Language

inclusionAI Adds Vision to Ling 3.0 Flash

The new Ling-3.0-flash-VL brings a mixture-of-experts vision-language model to inclusionAI's open lineup under a permissive MIT license.

Sep 4, 2026
Text / LLMVision-Language
Ling-3.0-flash-VL
Unknown/Text / LLM

Occamy-1.0 targets long-horizon agent work at 35B

A new open 35B model is tuned for multi-step, co-working tasks rather than one-shot answers.

Sep 3, 2026
Text / LLMReasoning
inclusionAI/Text / LLM

inclusionAI Tunes Ling-3.0-flash for Finance

A finance-focused variant of the Ling-3.0-flash MoE model targets financial research and agentic tool use.

Sep 3, 2026
Text / LLM
Ling-3.0-flash-Fin
Fla Hub/Text / LLM

RWKV7-G1j arrives as a 13.3B attention-free model

The latest RWKV7 checkpoint scales the recurrent, attention-free architecture to 13.3 billion parameters under a permissive Apache 2.0 license.

Sep 2, 2026
Text / LLM
RWKV7-G1j-13.3B
IFM/Text / LLM

IFM releases K2-Horizon, a 375B open-weight MoE

The flagship model uses a mixture-of-experts design that activates just 23 billion parameters per token, keeping inference costs in check.

Sep 1, 2026
Text / LLMReasoning
K2-Horizon-375B-A23B
IFM/Text / LLM

IFM releases K2-Horizon-7B with open pretraining data

The dense 7-billion-parameter model ships as open weights alongside the datasets used to train it.

Sep 1, 2026
Text / LLM
K2-Horizon-375B-A23B
IFM/Text / LLM

IFM's K2-Horizon MoVA Ships as a 36B MoE Model

The open-weight language model activates just 4B of its 36B parameters per token, aiming for efficiency without shedding capacity.

Sep 1, 2026
Text / LLM
K2-Horizon-MoVA-36B-A4B
DeepSeek/Vision-Language

DeepSeek adds vision to its V4 Flash line

An experimental, MIT-licensed vision-language model brings image understanding to DeepSeek's fast V4 Flash architecture.

Aug 31, 2026
Text / LLMVision-Language
DeepSeek-V4-Flash-Vision-Exp
Tencent/Text / LLM

Tencent Previews Hunyuan Hy4, an Apache MoE Model

The company's next-generation Hunyuan language model arrives as an early preview with a permissive license and a mixture-of-experts design.

Aug 27, 2026
Text / LLMReasoning
Hunyuan Hy4 (preview)
IBM/Text / LLM

IBM's Granite 4.2 Adds Reasoning to Open LLM Line

The latest update to IBM's Apache 2.0 model family leans into structured reasoning while keeping its enterprise-friendly licensing.

Aug 25, 2026
Text / LLMReasoning
Zhipu AI/Text / LLM

Zhipu releases GLM-5.3-Flash under MIT license

A speed-tuned member of the GLM-5.3 family arrives with open weights and mixture-of-experts design aimed at fast, low-cost inference.

Aug 25, 2026
Text / LLMReasoning
GLM-5.3-Flash
Pipecat Ai/Text / LLM

NVIDIA's PhoneLLM Targets Voice Agents on the Line

An early alpha release brings a Nemotron-H mixture-of-experts model tuned for phone-based tool use and function calling.

Aug 24, 2026
Text / LLM
phonellm-alpha-1
XHToken/Text / LLM

XHToken releases Spark-X2.5, a 4B open LLM

The Apache-2.0 instruction-tuned model targets developers who want a small, permissively licensed text generator.

Aug 24, 2026
Text / LLM
Spark-X2.5-4B
Zhipu AI/Text / LLM

Zhipu's GLM-5.3 Targets Coding at a Fraction of the Cost

The open-weight MoE model from Zhipu AI aims to match frontier closed systems on coding tasks while undercutting them on price.

Aug 23, 2026
CodeText / LLM
Unknown/Text / LLM

Liquid AI's LFM2.5-DSpark targets faster inference

The company's latest LFM2.5 variant promises up to 3.2x faster inference without leaning on cloud-scale hardware.

Aug 20, 2026
Text / LLM
NVIDIA/Text / LLM

Liquid AI's LFM2.5-DSpark targets faster inference

The new efficient language model claims up to 3.2x faster inference, extending Liquid AI's push toward lean, deployable models.

Aug 20, 2026
Text / LLM
Thomsonreuters/Text / LLM

Thomson Reuters enters the model race with Thomson-1.0

The information giant's first frontier model is a mixture-of-experts system tuned on its proprietary legal, tax, and news data.

Aug 18, 2026
Text / LLMVision-Language
Thomson-1.0-Small
Ornith Ai/Text / LLM

Ornith 1.5 arrives as a 397B MoE multimodal model

The MIT-licensed release spans a 397B mixture-of-experts flagship plus 9B and 35B-A3B variants for lighter deployments.

Aug 18, 2026
Text / LLMVision-Language
Ornith-1.5-397B
Ornith Ai/Text / LLM

Ornith 1.5 Brings a Lean 35B MoE to Open Weights

The MIT-licensed model activates just 3B parameters per token, borrowing the Qwen3.5 MoE design for text and reasoning tasks.

Aug 18, 2026
Text / LLMReasoning
Ornith-1.5-397B
Apodex/Text / LLM

Apodex 1.1 mini targets long-horizon agentic work

An MoE model built on Qwen3.5-35B-A3B aims at complex, multi-step tasks rather than one-shot answers.

Aug 17, 2026
Text / LLMReasoning
Apodex 1.1 mini
OpenMOSS/Vision-Language

OpenMOSS Debuts MOSS-VL for Real-Time Vision Interaction

A new open vision-language model family uses gated cross-attention to enable streaming, low-latency multimodal exchanges.

Aug 14, 2026
Text / LLMVision-Language
DeepSeek/Text / LLM

DeepSeek Releases V4-Pro-0813 With Open Weights

The Chinese lab pushes a higher-capability checkpoint of its V4 line to Hugging Face under a permissive MIT license.

Aug 13, 2026
Text / LLMReasoning
DeepSeek-V4-Pro-0813
DeepSeek/Text / LLMMajor release

DeepSeek Releases V4-Pro, an MIT-Licensed MoE Model

The company's newest flagship targets reasoning and coding while keeping a permissive open-source license.

Aug 13, 2026
CodeText / LLM
DeepSeek-V4-Pro-0813
inclusionAI/Text / LLM

inclusionAI Releases Ling-3.0-tiny, an MIT-Licensed MoE

The new hybrid mixture-of-experts model targets efficient text generation under a permissive license.

Aug 10, 2026
Text / LLM
Ling-3.0-tiny
Qwen · Alibaba/Text / LLMMajor release

Qwen releases 2.4T-parameter open MoE with 95B active

Alibaba's Qwen team pushes its largest sparse model yet, activating 95 billion parameters per token from a 2.4-trillion-parameter pool.

Aug 8, 2026
Text / LLMReasoning
Qwen3.8-2.4T-A95B
IBM/Text / LLM

IBM's Granite 4.2 30B adds reasoning and tool calls

The new dense model brings optional step-by-step thinking and function calling under a permissive Apache-2.0 license.

Aug 7, 2026
Text / LLMReasoning
Granite 4.2 30B
Motif Technologies/Text / LLM

Motif 3 brings a sparse MoE approach to reasoning

Motif Technologies debuts a mixture-of-experts language model built around grouped differential latent attention for long-context reasoning and code.

Aug 7, 2026
CodeText / LLM
Motif 3
OpenMOSS/Text / LLM

MameLoshnLM Brings Yiddish Into the Open LLM Era

OpenMOSS releases what it calls the first open-source 8B language model built specifically for Yiddish, paired with an evaluation benchmark.

Aug 5, 2026
Text / LLM
Qwen · Alibaba/Vision-LanguageMajor release

Qwen3.8-27B Brings Vision to a Dense Model

Alibaba's latest Qwen release pairs image understanding with text under a permissive Apache-2.0 license.

Aug 5, 2026
Text / LLMVision-Language
Qwen3.8-27B
Deepgrove/Reasoning

Deepgrove's Maple Preview bets on ternary-weight MoE

A permissively licensed reasoning model that pairs a mixture-of-experts design with ternary weights, aiming for efficiency.

Aug 4, 2026
Text / LLMReasoning
Maple Preview
NVIDIA/Text / LLM

NVIDIA's Nemotron 3.5 Lightning trims MoE for speed

A 30-billion-parameter mixture-of-experts model activates just 3 billion parameters per token, using a hybrid Mamba design to keep inference fast.

Aug 4, 2026
Text / LLMReasoning
Nemotron-3.5-Lightning-30B-A3B
inclusionAI/Text / LLM

inclusionAI ships Ling-3.0-flash, an MIT-licensed MoE model

The latest entry in the Bailing family pairs a hybrid mixture-of-experts design with a permissive license aimed at fast, low-cost text generation.

Aug 2, 2026
Text / LLM
Ling-3.0-flash
LiquidAI/Embeddings

Liquid AI ships LFM2.5, a 2.6B on-device model

The latest Liquid Foundation Model targets multilingual text generation on laptops and phones, packaged in GGUF for local runtimes.

Aug 1, 2026
Text / LLM
LFM2.5-2.6B
NVIDIA/Text / LLM

NVIDIA's Nemotron 3.5 Lightning Blends Mamba and MoE

A 30B mixture-of-experts model with just 3B active parameters aims at fast, agentic coding workloads.

Aug 1, 2026
Text / LLMReasoning
Nemotron-3.5-Lightning-30B-A3B
Meituan/Text / LLM

Meituan Ships a Lighter, Sparser LongCat-Flash

The food-delivery giant's newest open model trims its mixture-of-experts design for more efficient inference under an MIT license.

Jul 31, 2026
Text / LLM
LongCat-Flash-Lite-Sparse
DeepSeek/Text / LLM

DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents

The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.

Jul 31, 2026
Text / LLMReasoning
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLMMajor release

DeepSeek Refreshes V4-Flash With New 0731 Checkpoint

The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.

Jul 31, 2026
Text / LLMReasoning
DeepSeek-V4-Flash-0731
Cactus Compute/Text / LLM

Needle2 Packs an Agentic LLM Into 14MB

Cactus Compute's tiny model brings tool and function calling to phones, wearables, and robots.

Jul 29, 2026
CodeText / LLM
Needle2
LGAI EXAONE/Text / LLM

LG AI Research debuts K-EXAONE 2.0, a 750B MoE model

The new mixture-of-experts model activates 37B parameters per token and targets English, Korean, and Spanish reasoning tasks.

Jul 29, 2026
Text / LLMReasoning
K-EXAONE 2.0 750B-A37B
LiquidAI/Embeddings

Liquid AI's LFM2.5-2.6B targets on-device agents

A compact 2.6-billion-parameter model built to run local, multilingual agents without leaning on the cloud.

Jul 28, 2026
Text / LLM
LFM2.5-2.6B
Skt/Text / LLM

SK Telecom Releases A.X-K2 Multilingual LLM

The Korean telecom carrier's latest open language model targets English, Korean, Chinese, Japanese, and Spanish under a permissive license.

Jul 28, 2026
Text / LLM
A.X-K2
Thinkingmachines/Vision-Language

Thinking Machines Debuts Inkling Small, a Compact Multimodal MoE

The Apache-2.0 model brings mixture-of-experts efficiency to image, audio, and text tasks in a smaller footprint.

Jul 27, 2026
Text / LLMAny-to-Any
Inkling Small
Swiss Ai/Text / LLM

Apertus v1.5 70B arrives with an Apache-2.0 license

Switzerland's open-model effort ships a 70-billion-parameter, multilingual and multimodal system that anyone can use, modify, and deploy.

Jul 24, 2026
Text / LLMVision-Language
Apertus v1.5 70B
Amd/Reasoning

AMD's Instella-MoE Brings Reasoning to ROCm Hardware

A new open mixture-of-experts model with 16B total parameters and just 3B active is tuned to run on AMD's own accelerator stack.

Jul 23, 2026
Text / LLMReasoning
Instella-MoE 16B A3B Think
Kwaipilot/Code

Kwaipilot Releases KAT-Coder V2.5 Dev, an Agentic MoE Coder

Kuaishou's coding team ships an open mixture-of-experts model built on the Qwen3.5 MoE architecture and tuned for agentic development work.

Jul 23, 2026
CodeText / LLM
KAT-Coder V2.5 Dev
Upstage/Text / LLM

Upstage's Solar Open2 arrives as a 250B MoE model

The Korean AI firm's latest open release scales to 250 billion parameters with a mixture-of-experts design tuned for English and Korean.

Jul 22, 2026
Text / LLM
Solar Open2 250B
Nanbeige/Text / LLM

Nanbeige 4.2 arrives as a compact 3B bilingual model

The new English-Chinese language model targets efficient deployment in a small parameter footprint.

Jul 21, 2026
Text / LLM
Nanbeige4.2-3B
Motif Technologies/Text / LLM

Motif Technologies debuts Motif 3 Beta, an MoE model

The Korean AI lab's preview release is a mixture-of-experts language model built for long-context, multilingual work.

Jul 20, 2026
Text / LLM
Motif 3
Unknown/Text / LLM

German consortium releases open 30B model Soofi S

A collaborative European effort ships a dense 30-billion-parameter model that claims top marks on both English and German benchmarks.

Jul 16, 2026
Text / LLM
Mistral AI/Vision-Language

Mistral's Shieldstral brings safety checks to images

A compact 3B open-weights classifier flags unsafe text and visual content, and ships under Apache 2.0.

Jul 16, 2026
Text / LLMVision-Language
Shieldstral 1.0 3B
inclusionAI/Text / LLM

inclusionAI ships LLaDA2.2-flash diffusion LLM

A new Apache-2.0 mixture-of-experts model that generates text through diffusion rather than left-to-right decoding.

Jul 16, 2026
Text / LLM
LLaDA2.2-flash
OpenAI/Any-to-Any

Thinking Machines ships Inkling, its first open model

The Mira Murati-founded lab makes its debut with an open-weights, reasoning-focused language model.

Jul 15, 2026
Text / LLMReasoning
Thinkingmachines/Any-to-AnyMajor release

Thinking Machines Lab debuts Inkling, its first open model

The lab's inaugural open-weights release is a mixture-of-experts system that takes image and audio inputs, shipped under a permissive Apache 2.0 license.

Jul 15, 2026
Text / LLMAny-to-Any
Inkling
inclusionAI/Reasoning

inclusionAI's Ring-Zero Scales Zero-RL to a Trillion Parameters

A new mixture-of-experts model learns to reason through reinforcement learning alone, without human-annotated chains of thought.

Jul 13, 2026
Text / LLMReasoning
Poolside/Code

Poolside releases Laguna-S-2.1 coding model

The AI coding startup puts a version of its Laguna family on Hugging Face under the permissive OpenMDW license.

Jul 13, 2026
CodeText / LLM
Laguna-S-2.1
Ai Sage/Text / LLM

GigaChat 3.5 arrives as a 432B mixture-of-experts model

The multilingual instruct model activates 28B parameters per token and leans on hybrid attention for efficiency at scale.

Jul 5, 2026
Text / LLM
GigaChat3.5-432B-A28B
Prism Ml/Text / LLM

Bonsai-27B Brings 1-Bit Quantization to Local Inference

A ternary-weight 27B model with hybrid attention aims to run large-model reasoning on everyday hardware.

Jul 4, 2026
Text / LLM
Bonsai-27B
Tencent/Text / LLM

Tencent releases Hunyuan Hy3 under Apache 2.0

The company's latest mixture-of-experts model arrives as an openly licensed conversational LLM on Hugging Face.

Jul 2, 2026
Text / LLM
Hunyuan Hy3
Google DeepMind/Any-to-AnyMajor release

Google DeepMind's Gemma 4 Goes Multimodal and MoE

The new open-weights family adds a mixture-of-experts design, encoder-free multimodal inputs, and an optional thinking mode.

Jul 1, 2026
Text / LLMAny-to-Any
Mistral AI/Text / LLM

Mistral's Leanstral 1.5 puts 119B in a lean MoE

The new Apache-2.0 mixture-of-experts model activates just 6B parameters per token, trading raw density for cheaper inference.

Jul 1, 2026
Text / LLMReasoning
Leanstral-1.5-119B-A6B
Soofi Project/Text / LLM

German Consortium Debuts Soofi S, an Open 30B MoE Model

A Mamba-2 mixture-of-experts model claims top marks in both English and German benchmarks.

Jul 1, 2026
Text / LLM
Soofi S Base
Mistral AI/Text / LLM

Liquid AI's LFM2.5-230M targets phones and robots

A 230-million-parameter model built to run on constrained hardware like Raspberry Pi and edge robotics.

Jul 1, 2026
Text / LLM
IBM/Text / LLM

Liquid AI's LFM2.5 230M targets phones and robots

A 230-million-parameter language model built to run locally on constrained hardware like the Raspberry Pi.

Jul 1, 2026
Text / LLM
NVIDIA/Text / LLM

Liquid AI's LFM2.5-230M targets phones and robots

A 230-million-parameter language model built to run on hardware as modest as a Raspberry Pi.

Jul 1, 2026
Text / LLM
Meituan/Text / LLM

Meituan releases LongCat-2.0 language model

The Chinese delivery giant continues its push into open AI with a new text model on Hugging Face.

Jun 30, 2026
Text / LLM
LongCat-2.0
DeepSeek/Text / LLMMajor release

DeepSeek Releases V4-Pro, an MIT-Licensed MoE Model

The new flagship arrives as a mixture-of-experts system with FP8 weights and open reasoning capabilities under a permissive license.

Jun 27, 2026
Text / LLMReasoning
DeepSeek-V4-Pro
DeepSeek/Text / LLM

DeepSeek Releases V4-Flash for Low-Latency Inference

A lighter, faster member of DeepSeek's V4 line arrives on Hugging Face under a permissive MIT license.

Jun 27, 2026
Text / LLM
DeepSeek V4.1 Flash
Deepreinforce Ai/Text / LLM

DeepReinforce Releases Ornith 1.0, a 35B Reasoning Model

The new dense model ships in GGUF format under a permissive MIT license, aimed at local and self-hosted deployment.

Jun 25, 2026
Text / LLMReasoning
Ornith-1.5-397B
LiquidAI/Text / LLM

Liquid AI's LFM2.5-230M targets on-device language tasks

A 230-million-parameter multilingual model built to run efficiently at the edge rather than in the cloud.

Jun 24, 2026
Text / LLM
LFM2.5 230M
NVIDIA/Text / LLM

NVIDIA's Nemotron 3 Puzzle Runs Big on a Lean Budget

A 75-billion-parameter mixture-of-experts reasoning model that activates just 9 billion parameters per token.

Jun 24, 2026
Text / LLMReasoning
Nemotron-Labs-3-Puzzle-75B-A9B
NVIDIA/Text / LLM

NVIDIA's Nemotron-3 Puzzle Brings a Lean MoE to Reasoning

The 75B-parameter model activates just 9B per token and ships in NVIDIA's NVFP4 format for efficient inference.

Jun 24, 2026
Text / LLMReasoning
Nemotron-3 Puzzle 75B-A9B
Deepreinforce Ai/Text / LLM

DeepReinforce debuts Ornith-1.0, a 397B MoE model

The flagship of a new open model family arrives under a permissive MIT license, with reasoning among its stated strengths.

Jun 23, 2026
Text / LLMReasoning
Ornith-1.5-397B
Qwen · Alibaba/Text / LLM

Qwen's AgentWorld Simulates Worlds for AI Agents

Alibaba's new MoE model acts as a language world model, generating the environments that agents act within.

Jun 22, 2026
Text / LLMReasoning
Qwen-AgentWorld-35B-A3B
InternScience/Reasoning

Agents-A1: A 35B MoE Built for Agentic Scaling

InclusionAI's new mixture-of-experts model bets that agent-horizon scaling can rival far larger systems on long-running tasks.

Jun 22, 2026
Text / LLMReasoning
Agents-A1
Deepreinforce Ai/Text / LLM

DeepReinforce's Ornith-1.0-9B Targets Agentic Coding

A compact, MIT-licensed 9B model built for autonomous coding tasks arrives on Hugging Face.

Jun 21, 2026
CodeText / LLM
Ornith-1.5-397B
Deepreinforce Ai/Text / LLM

Ornith-1.0-35B brings a mid-size MoE to agentic coding

An MIT-licensed mixture-of-experts model targets self-scaffolding code tasks without the footprint of a frontier system.

Jun 21, 2026
CodeText / LLM
Ornith-1.5-397B
Poolside/Code

Poolside releases Laguna XS 2.1 code model

The compact, code-focused language model arrives on Hugging Face under an open model license.

Jun 20, 2026
CodeText / LLM
Laguna-XS-2.1
Zhipu AI/Text / LLMMajor release

Zhipu AI Releases MIT-Licensed GLM-5.2 MoE Model

The new bilingual model from the Chinese AI firm uses a Mixture of Experts architecture and sparse attention under a fully permissive license.

Jun 17, 2026
Text / LLMReasoning
GLM-5.3
Poolside/Text / LLM

Poolside Releases Laguna-M.1, an Open MoE Model

The AI coding startup steps into open weights with an Apache-2.0 mixture-of-experts model built for text and code.

Jun 15, 2026
CodeText / LLM
Laguna-M.1
Microsoft/Text / LLM

Microsoft's FastContext is a 4B sub-agent for code

A compact Qwen3-derived model built to explore repositories, released under a permissive MIT license.

Jun 14, 2026
CodeText / LLM
FastContext 1.0 4B SFT
Moonshot AI/Text / LLMMajor release

Moonshot AI releases Kimi K3, a 2.8T-parameter MoE model

The open-weights multimodal model leans into coding and agentic tasks, extending Moonshot's Kimi line into a new scale bracket.

Jun 13, 2026
Text / LLMReasoning
Kimi K3
WeiboAI/Reasoning

Weibo AI Releases VibeThinker-3B, a Compact Reasoning Model

The new 3-billion-parameter model from the Chinese tech giant focuses on challenging benchmarks in mathematics, coding, and graduate-level questions.

Jun 12, 2026
CodeText / LLM
VibeThinker-3B
Moonshot AI/CodeMajor release

Moonshot AI Releases Kimi, a Multimodal Coding Model

The new Mixture-of-Experts model from the Chinese AI company can generate code while also understanding visual inputs, a rare combination in open models.

Jun 11, 2026
CodeText / LLM
Kimi-K2.7-Code
Google DeepMind/Text / LLM

Google Releases Open-Source DiffusionGemma 26B Model

The new 26B parameter model from DeepMind uses a diffusion-based architecture, a technique more common in image generation, to produce text.

Jun 9, 2026
Text / LLMVision-Language
DiffusionGemma
Cohere/Code

Cohere Releases North-Mini-Code, an Open MoE Model

The new Apache 2.0-licensed model is designed for code generation and agentic chat applications, using a Mixture-of-Experts architecture for efficiency.

Jun 5, 2026
CodeText / LLM
North-Mini-Code 1.0
Google DeepMind/Any-to-AnyMajor release

Google Releases Gemma 4 12B Multimodal Model

The new 12-billion-parameter open model from DeepMind introduces a unified 'any-to-any' architecture for advanced multimodal tasks.

May 23, 2026
Text / LLMAny-to-Any
Gemma 4
Google DeepMind/Any-to-AnyMajor release

Google Releases Gemma 4, a 12B 'Any-to-Any' Model

The new 12-billion-parameter model from Google DeepMind is designed to handle a flexible mix of data types, moving beyond traditional text and image inputs.

May 23, 2026
Text / LLMAny-to-Any
Gemma 4
OpenBMB/Text / LLM

OpenBMB's MiniCPM5-1B targets on-device AI

The compact 1B-parameter model brings long-context handling and tool-calling to phones and laptops.

May 21, 2026
Text / LLM
MiniCPM5-2B
OpenBMB/Text / LLM

GLiGuard: A Sub-1B Model for Faster LLM Guardrails

The team behind GLiNER releases an open-source small language model aimed at making safety moderation cheaper and quicker to run.

May 12, 2026
Text / LLM
OpenBMB/Code

Needle: A 26M-Parameter Model Built for Tool Calling

Cactus Compute distilled Gemini's tool-calling behavior into a tiny model meant to run locally.

May 12, 2026
CodeText / LLM
Tencent/Text / LLM

Tencent Releases 1.8B Model for Multilingual Translation

The 1.8 billion-parameter model from the Chinese tech giant is designed for high-quality translation across a wide range of language pairs.

May 11, 2026
Text / LLM
Hunyuan-MT2 1.8B
Moonshot AI/Vision-LanguageMajor release

Kimi K2.6 tops closed models in coding test

Moonshot AI's open-weights mixture-of-experts model reportedly outperformed Claude, GPT-5.5, and Gemini on a programming challenge.

May 3, 2026
CodeText / LLM
Google DeepMind/Any-to-Any

Google Releases Gemma 4 Multimodal Open Model

The new 26-billion-parameter model from DeepMind uses a mixture-of-experts design for greater efficiency and is tuned for assistant-style tasks.

Apr 23, 2026
Text / LLMAny-to-Any
Gemma 4 26B-A4B Instruct (MoE)
Google DeepMind/Any-to-AnyMajor release

Google Releases Multimodal Gemma 4 31B Model

The new 31-billion-parameter model is an instruction-tuned, 'any-to-any' powerhouse released under a permissive Apache 2.0 license.

Apr 23, 2026
Text / LLMAny-to-Any
Gemma 4
Google DeepMind/Any-to-Any

Google Releases 4B Multimodal Gemma 4 Assistant

The new 4-billion-parameter model is instruction-tuned for 'any-to-any' tasks, handling a flexible mix of data types.

Apr 23, 2026
Text / LLMAny-to-Any
Gemma 4 E4B-it Assistant
Google DeepMind/Any-to-Any

Google Releases 2B Multimodal Gemma 4 Assistant Model

The new compact model from DeepMind is instruction-tuned for "any-to-any" tasks, capable of processing and generating mixed data types.

Apr 23, 2026
Text / LLMAny-to-Any
Gemma 4 E2B-it Assistant
DeepSeek/Text / LLMMajor release

DeepSeek Releases V4-Pro, an Open MoE Contender

The new flagship model combines a Mixture-of-Experts architecture with a permissive MIT license, positioning it for wide commercial adoption.

Apr 22, 2026
CodeText / LLM
DeepSeek-V4-Pro
DeepSeek/Text / LLMMajor release

DeepSeek Releases V4-Flash, a Fast MIT-Licensed MoE Model

The new Mixture of Experts model from the Beijing-based AI lab is optimized for fast, efficient conversational AI and carries a fully permissive license.

Apr 22, 2026
Text / LLMReasoning
DeepSeek V4.1 Flash
Qwen · Alibaba/Vision-Language

Alibaba's Qwen Releases Open 27B Vision Model

The new dense model, licensed under Apache 2.0, brings both text and image understanding to the midrange parameter space.

Apr 21, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-LanguageMajor release

Qwen Releases 35B Multimodal Mixture-of-Experts Model

The new Qwen3.6-35B-A3B from Alibaba's Qwen team combines vision and language capabilities using an efficient sparse architecture.

Apr 15, 2026
Text / LLMReasoning
Qwen3.8-27B
Moonshot AI/Vision-LanguageMajor release

Moonshot AI Releases Kimi-K2.6 Multimodal Model

The Chinese AI lab has published weights for its new vision-language model, though a restrictive license limits its use to research applications.

Apr 14, 2026
Text / LLMVision-Language
Kimi K2.6
NVIDIA/Text / LLM

NVIDIA's Nemotron TwoTower mixes diffusion and Mamba

A new 30B mixture-of-experts base model activates just 3B parameters per token and pairs a hybrid diffusion/Mamba design.

Apr 11, 2026
Text / LLM
Nemotron TwoTower 30B-A3B Base
NVIDIA/Text / LLM

NVIDIA's Nemotron TwoTower is a MoE experiment

An experimental 30B mixture-of-experts base model blends diffusion and Mamba ideas under a two-tower design.

Apr 11, 2026
Text / LLM
Nemotron TwoTower 30B-A3B Base
MiniMax/Text / LLM

MiniMax Releases M2.7, an MoE Model with FP8 Weights

The new conversational language model from the Chinese AI company uses a Mixture-of-Experts architecture and 8-bit weights, but is released under a restrictive custom license.

Apr 9, 2026
Text / LLMReasoning
MiniMax-M2.7
Zhipu AI/Text / LLMMajor release

Zhipu AI Releases Open-Source GLM-5.1 MoE Model

The new bilingual model from the Chinese AI firm features an efficient Mixture-of-Experts architecture and a fully permissive MIT license.

Apr 3, 2026
Text / LLMReasoning
GLM-5.3
Meituan/Any-to-Any

Meituan Releases LongCat-Next 'Any-to-Any' AI Model

The Chinese tech company has released the weights for a unified model that can process and generate combinations of text, images, audio, and video.

Mar 25, 2026
Text / LLMAny-to-Any
LongCat-Next
Cactus Compute/Code

Needle: A 26M-Param Model Built for On-Device Tool Calls

Cactus Compute's tiny encoder-decoder is distilled specifically for function calling at the edge, trading general chat for a narrow, useful job.

Mar 16, 2026
CodeText / LLM
Needle
Google DeepMind/Any-to-AnyMajor release

Google Releases Gemma 4, a 26B Vision-Language Model

The new open-source model from DeepMind uses a Mixture-of-Experts architecture to handle both text and image inputs efficiently.

Mar 11, 2026
Text / LLMVision-Language
Gemma 4
Google DeepMind/Any-to-AnyMajor release

Google Releases Multimodal Gemma 4 31B Model

The new 31-billion-parameter model is instruction-tuned and can process both text and images, marking a significant expansion for the Gemma family.

Mar 11, 2026
Text / LLMVision-Language
Gemma 4
Google DeepMind/Any-to-AnyMajor release

Google Releases Compact Gemma 4 E2B Multimodal Model

The new 2-billion-parameter model from Google DeepMind brings efficient image-and-text understanding to the open-source Gemma family.

Mar 2, 2026
Text / LLMAny-to-Any
Gemma 4 E2B
Google DeepMind/Any-to-AnyMajor release

Google's Gemma 4 Arrives with Any-to-Any Multimodal Skills

The new 2-billion-parameter model from DeepMind can process text, vision, and audio, making it a versatile and efficient foundation for developers.

Mar 2, 2026
Text / LLMAny-to-Any
Gemma 4 E2B
Google DeepMind/Any-to-Any

Google Releases Gemma 4 E4B, a 4B Multimodal Model

The new 4-billion-parameter vision-language model brings image and text understanding to Google's popular open-source family.

Mar 2, 2026
Text / LLMAny-to-Any
Gemma 4 E4B
Google DeepMind/Any-to-AnyMajor release

Google's Gemma 4 Debuts with Any-to-Any Multimodality

The new 4-billion parameter model from Google DeepMind is designed for versatile input and output, handling text, images, and other data types.

Mar 2, 2026
Text / LLMAny-to-Any
Gemma 4 E4B
Qwen · Alibaba/Vision-Language

Alibaba's Qwen Releases Compact 0.8B Vision Model

The new 800-million-parameter model is the smallest in the Qwen3.5 family, designed for efficient multimodal tasks on consumer-grade hardware.

Feb 28, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-Language

Alibaba's Qwen team releases 4B vision-language model

The new Qwen3.5-4B model combines text and image understanding in a compact, permissively licensed package for developers.

Feb 27, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-Language

Qwen Releases 9B Multimodal Model in New 3.5 Series

The new open-source vision-language model from Alibaba's Qwen team offers strong performance in a compact, Apache 2.0-licensed package.

Feb 27, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-LanguageMajor release

Qwen Releases Flagship 122B Multimodal MoE Model

The new Qwen3.5-122B-A10B combines a massive parameter count with an efficient Mixture-of-Experts architecture for advanced vision and language tasks.

Feb 24, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-LanguageMajor release

Qwen Releases 27B Vision Model with Long Context

The new model from Alibaba's Qwen team combines multimodal understanding with a 131K token context window under a permissive Apache 2.0 license.

Feb 24, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-Language

Qwen Releases Efficient 35B Multimodal MoE Model

The new Qwen3.5-35B-A3B model from Alibaba combines vision and language capabilities with a resource-friendly Mixture of Experts design.

Feb 24, 2026
Text / LLMVision-Language
Qwen3.8-27B
Qwen · Alibaba/Vision-LanguageMajor release

Qwen releases flagship 397B multimodal MoE

The new open-source model from Alibaba uses a Mixture-of-Experts architecture to balance massive scale with efficient inference.

Feb 16, 2026
Text / LLMVision-Language
Qwen3.8-27B
MiniMax/Text / LLM

MiniMax Releases M2.5 Mixture-of-Experts Model

The Chinese AI company's first open-weight release uses an efficient FP8 data type but comes with a restrictive, non-commercial license.

Feb 12, 2026
Text / LLM
MiniMax-M2.7
Zhipu AI/Text / LLMMajor release

Zhipu AI Releases Open-Source GLM-5 MoE Model

The new Mixture-of-Experts model from the Chinese AI company combines an advanced architecture with a fully permissive MIT license for commercial use.

Feb 11, 2026
Text / LLMReasoning
GLM-5.3
Nanbeige/Text / LLM

Nanbeige Releases 3B Chinese-Enhanced Language Model

The new Llama-based model was trained from scratch on 3.5 trillion tokens of Chinese and English data to enhance its bilingual capabilities.

Feb 10, 2026
Text / LLM
Nanbeige4.2-3B
Qwen · Alibaba/Code

Qwen Releases Coder-Next, A New Open MoE Coding Model

The new model from Alibaba's Qwen team uses a Mixture-of-Experts architecture and is released under the commercially-friendly Apache 2.0 license.

Jan 30, 2026
CodeText / LLM
Qwen3-Coder-Next
Zhipu AI/Text / LLM

Zhipu AI Releases GLM-4.7-Flash MoE Model

The new Mixture-of-Experts model from the Beijing-based AI company is optimized for speed and released under the permissive MIT license.

Jan 19, 2026
Text / LLM
GLM-4.7-Flash
Google DeepMind/Text / LLM

Google Releases TranslateGemma for Open Translation

The new 4B-parameter model is an instruction-tuned variant of Gemma, designed specifically for high-quality multilingual translation tasks.

Jan 12, 2026
Text / LLM
TranslateGemma 4B IT
Moonshot AI/Vision-LanguageMajor release

Moonshot AI Releases Kimi K2.5 Multimodal Model

The new vision-language model from the Chinese AI firm uses a Mixture-of-Experts architecture and is now available on Hugging Face.

Jan 1, 2026
Text / LLMReasoning
Kimi K2.6
MiniMax/Text / LLM

MiniMax Debuts M2.1, an MoE Model Optimized with FP8

The new Mixture of Experts model from the Chinese AI firm uses 8-bit floating-point precision for a smaller memory footprint and faster inference.

Dec 20, 2025
Text / LLM
MiniMax-M2.7
DeepSeek/Text / LLM

DeepSeek-V3.2 Arrives With FP8 Weights, MIT License

The new Mixture-of-Experts model from DeepSeek AI combines an efficient FP8 architecture with a fully permissive license for commercial use.

Dec 1, 2025
Text / LLM
DeepSeek-V3.2
Moonshot AI/ReasoningMajor release

Moonshot AI Releases Kimi-K2 Reasoning Model

The new Mixture-of-Experts model is designed for complex tasks but arrives in a custom compressed format with a restrictive license.

Nov 4, 2025
Text / LLMReasoning
Kimi K2 Thinking
MiniMax/Text / LLMMajor release

MiniMax Releases M2, an Open-Weight MoE for Agents

The Shanghai-based AI startup has released a new Mixture-of-Experts model focused on complex reasoning, coding, and agentic tasks.

Oct 22, 2025
CodeText / LLM
MiniMax-M2.7
Google DeepMind/Text / LLM

Google Releases Compact FunctionGemma Model

The new 270-million-parameter model from Google DeepMind is fine-tuned specifically for reliable function calling and tool use.

Oct 8, 2025
Text / LLM
FunctionGemma 270M IT
Zhipu AI/Text / LLMMajor release

Zhipu AI Releases Open-Weight MoE Model GLM-4.6

The new Mixture-of-Experts model is available under a permissive MIT license and is optimized for complex reasoning and coding tasks.

Sep 29, 2025
Text / LLMReasoning
GLM-4.6
Qwen · Alibaba/Text / LLMMajor release

Qwen Releases 80B Mixture-of-Experts Model

The new Qwen3-Next model from Alibaba combines a large parameter count with an efficient MoE architecture to balance performance and computational cost.

Sep 9, 2025
Text / LLM
Qwen3-Next-80B-A3B-Instruct
DeepSeek/Text / LLMMajor release

DeepSeek Releases 671B MoE Model Under MIT License

The new DeepSeek-V3.1-Base is a massive 671-billion-parameter Mixture-of-Experts model designed for efficient, large-scale research and development.

Aug 19, 2025
Text / LLMReasoning
DeepSeek-V3.2
Google DeepMind/Text / LLM

Google Releases Gemma 3 270M for On-Device AI

The new ultra-compact model from DeepMind is designed for efficient performance in resource-constrained environments like mobile and web.

Aug 5, 2025
Text / LLM
Gemma 3 270M
OpenAI/ReasoningMajor release

OpenAI Releases 21B Open-Weight MoE Model

The new `gpt-oss-20b` is an Apache 2.0-licensed Mixture-of-Experts model designed to run efficiently on consumer-grade hardware.

Aug 4, 2025
Text / LLMReasoning
gpt-oss-20b
OpenAI/ReasoningMajor release

OpenAI Releases Its First Open-Source MoE Model

The new 117-billion-parameter `gpt-oss-120b` is a Mixture-of-Experts model focused on reasoning, released under a permissive Apache 2.0 license.

Aug 4, 2025
Text / LLMReasoning
gpt-oss-20b
Qwen · Alibaba/Code

Qwen Releases Compact 30B MoE for Coding Agents

The new Apache 2.0 model from Alibaba's Qwen team uses a Mixture-of-Experts architecture to deliver strong performance with only 3B active parameters.

Jul 31, 2025
CodeText / LLM
Qwen3-Coder-30B-A3B-Instruct
Qwen · Alibaba/CodeMajor release

Qwen Releases 480B Open-Source Model for Code Agents

The new flagship coding model from Alibaba's Qwen team uses a massive Mixture-of-Experts architecture and is released under a permissive Apache-2.0 license.

Jul 22, 2025
CodeText / LLM
Qwen3-Coder-30B-A3B-Instruct
Zhipu AI/Text / LLMMajor release

Z.ai Releases 355B Parameter GLM-4.5 Under MIT License

The new Mixture-of-Experts model combines massive scale with a fully permissive license, targeting complex reasoning and agentic applications.

Jul 20, 2025
CodeText / LLM
GLM-4.6
Moonshot AI/Vision-LanguageMajor release

Moonshot AI Releases Trillion-Parameter Kimi-K2 Model

The new Mixture-of-Experts model brings massive scale to the open-weights community, focusing on complex reasoning and coding tasks with a 128K context window.

Jul 11, 2025
Text / LLMReasoning
Kimi K2.6