The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestZhipu AI4.7-Flash
Zhipu AIText / LLM

Zhipu AI Releases GLM-4.7-Flash MoE Model

The new Mixture-of-Experts model from the Beijing-based AI company is optimized for speed and released under the permissive MIT license.

Jan 19, 2026
NotableMIT
GLM-4.7-Flash

Chinese AI firm Zhipu AI has released GLM-4.7-Flash, a new language model designed for high-speed inference. It employs a Mixture-of-Experts (MoE) architecture, a technique that allows models to scale up their parameter counts while keeping computational costs manageable during inference.

The 'Flash' in its name signals the model's primary goal: performance. MoE models achieve this by selectively activating only a fraction of their total parameters—the 'experts'—to process any given input. This makes them significantly faster and more efficient for real-time applications compared to dense models of a similar size.

A Permissive License for Commercial Use

Perhaps most notably for developers and businesses, GLM-4.7-Flash is available under the MIT license. This is one of the most permissive open-source licenses, imposing very few restrictions on reuse and allowing for broad commercial applications. This combination of an efficient architecture and a business-friendly license makes the model an attractive option for integration into products and services.

The model is the latest addition to Zhipu AI's GLM-4 family of models. While specific details on its total parameter count and context length have not been disclosed, developers can access the model weights and resources on its official Hugging Face repository.

Sources

  • zai-org/GLM-4.7-Flash

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Active params3B active
Size62.4 GB
PrecisionBF16
ArchitectureGlm4MoeLiteForCausalLM
LicenseMIT
Downloads2.1M
Likes1.8K

Modalities

Text / LLM

0 comments

No comments yet. Be the first to weigh in.

More in Text / LLM

LongCat-Flash-Lite-Sparse
Meituan/Text / LLM

Meituan Ships a Lighter, Sparser LongCat-Flash

The food-delivery giant's newest open model trims its mixture-of-experts design for more efficient inference under an MIT license.

Jul 31, 2026
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Refreshes V4-Flash With New 0731 Checkpoint

The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.

Jul 31, 2026
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents

The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.

Jul 31, 2026