DeepSeek Releases V4-Pro, an MIT-Licensed MoE Model
The new flagship arrives as a mixture-of-experts system with FP8 weights and open reasoning capabilities under a permissive license.

DeepSeek has released DeepSeek-V4-Pro, the latest entry in its widely watched line of open models, publishing the weights on Hugging Face. The model is a mixture-of-experts (MoE) system shipped with FP8 weights and made available under the permissive MIT license.
The release is positioned as a text and reasoning model, continuing DeepSeek's push into systems that can handle both general language tasks and multi-step problem solving. As with the company's earlier work, distributing the model in FP8 format signals an emphasis on efficient inference, letting the large expert-based architecture run with a smaller memory footprint than full-precision equivalents.
Why it matters
DeepSeek has become one of the most closely tracked labs in open-weight AI, and each major version tends to reset expectations for what freely available models can do. A few points stand out here:
- MIT licensing places few restrictions on commercial and research use, a key draw for teams building on open models.
- Mixture-of-experts design activates only part of the network per token, a route to scaling capacity without proportional compute cost.
- FP8 weights lower the barrier to deployment on constrained hardware.
Detailed specifications such as parameter counts, context length, and benchmark results were not included in the release record, so builders will want to consult the official model card as DeepSeek fills in documentation. For now, the arrival of a new MIT-licensed flagship gives the open-source community another capable foundation to test, fine-tune, and deploy.
Sources
- Visit
deepseek-ai/DeepSeek-V4-Pro-DSpark
Hugging Face
More in Text / LLM
Agnes-3.0-Flash arrives as a multimodal reasoning model
The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

InternLM's Atria Dawn Preview Targets Agentic Tasks
A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.
ZGCM-1 arrives as a fully open 7B reasoning model
A compact foundation model targets math reasoning and agentic search with tool use, and its makers are releasing it fully open.
0 comments
No comments yet. Be the first to weigh in.