Zhipu AI Releases GLM-4.7-Flash MoE Model
The new Mixture-of-Experts model from the Beijing-based AI company is optimized for speed and released under the permissive MIT license.
Chinese AI firm Zhipu AI has released GLM-4.7-Flash, a new language model designed for high-speed inference. It employs a Mixture-of-Experts (MoE) architecture, a technique that allows models to scale up their parameter counts while keeping computational costs manageable during inference.
The 'Flash' in its name signals the model's primary goal: performance. MoE models achieve this by selectively activating only a fraction of their total parameters—the 'experts'—to process any given input. This makes them significantly faster and more efficient for real-time applications compared to dense models of a similar size.
A Permissive License for Commercial Use
Perhaps most notably for developers and businesses, GLM-4.7-Flash is available under the MIT license. This is one of the most permissive open-source licenses, imposing very few restrictions on reuse and allowing for broad commercial applications. This combination of an efficient architecture and a business-friendly license makes the model an attractive option for integration into products and services.
The model is the latest addition to Zhipu AI's GLM-4 family of models. While specific details on its total parameter count and context length have not been disclosed, developers can access the model weights and resources on its official Hugging Face repository.
Sources
- Visit
zai-org/GLM-4.7-Flash
Hugging Face
More in Text / LLM
Agnes-3.0-Flash arrives as a multimodal reasoning model
The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

InternLM's Atria Dawn Preview Targets Agentic Tasks
A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.
ZGCM-1 arrives as a fully open 7B reasoning model
A compact foundation model targets math reasoning and agentic search with tool use, and its makers are releasing it fully open.
0 comments
No comments yet. Be the first to weigh in.