Mistral Large 4 arrives as a trillion-param MoE
The new flagship is a sparse mixture-of-experts model with roughly 49B active parameters per token.
Mistral AI has introduced Mistral Large 4, the latest flagship in its Large series and the company's biggest model to date. According to the announcement, it is a mixture-of-experts (MoE) design with around one trillion total parameters but only about 49 billion active per token — a structure that aims to deliver the capacity of a very large model while keeping inference costs closer to a mid-sized one.
The model is positioned as a text and reasoning system, placing it among the growing class of frontier models tuned for multi-step problem solving rather than raw chat alone. Mistral is framing this release as a preview, so some details and availability specifics are still settling.
Why it matters
Sparse MoE architectures have become the dominant approach at the high end, and the 49B-active figure is the number worth watching. It signals that Mistral wants Large 4 to be deployable at a reasonable serving cost despite its trillion-parameter footprint.
- Scale: roughly 1T total parameters, a step up for the Large family
- Efficiency: about 49B active parameters per token via MoE routing
- Focus: text generation plus reasoning
Licensing terms for Large 4 fall outside the standard open templates, so teams evaluating it should read the usage conditions closely before committing. As Simon Willison noted in his writeup, the headline here is sheer size — but the practical story will depend on how the active-parameter economics hold up in real deployments.
Sources
- Visit
Introducing Mistral Large 4: Le chonk
Announcement
More in Text / LLM
Falcon-Emirati tunes an LLM for local dialect
TII's Falcon family gets a variant built around Emirati Arabic, aiming at culture and nuance rather than generic Gulf Arabic.
Reflection AI debuts Beam, a 501B open model
The startup's first open-weight release is a dense 501-billion-parameter model aimed at text and reasoning tasks.
Reflection releases Beam, a 501B open-weight model
The startup's first frontier-scale model ships with downloadable weights and a mixture-of-experts design aimed at reasoning.
0 comments
No comments yet. Be the first to weigh in.