The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDeepSeekV4
DeepSeekText / LLM

DeepSeek Releases V4-Pro, an MIT-Licensed MoE Model

The new flagship arrives as a mixture-of-experts system with FP8 weights and open reasoning capabilities under a permissive license.

Jun 27, 2026
Major releaseMIT
DeepSeek-V4-Pro

DeepSeek has released DeepSeek-V4-Pro, the latest entry in its widely watched line of open models, publishing the weights on Hugging Face. The model is a mixture-of-experts (MoE) system shipped with FP8 weights and made available under the permissive MIT license.

The release is positioned as a text and reasoning model, continuing DeepSeek's push into systems that can handle both general language tasks and multi-step problem solving. As with the company's earlier work, distributing the model in FP8 format signals an emphasis on efficient inference, letting the large expert-based architecture run with a smaller memory footprint than full-precision equivalents.

Why it matters

DeepSeek has become one of the most closely tracked labs in open-weight AI, and each major version tends to reset expectations for what freely available models can do. A few points stand out here:

  • MIT licensing places few restrictions on commercial and research use, a key draw for teams building on open models.
  • Mixture-of-experts design activates only part of the network per token, a route to scaling capacity without proportional compute cost.
  • FP8 weights lower the barrier to deployment on constrained hardware.

Detailed specifications such as parameter counts, context length, and benchmark results were not included in the release record, so builders will want to consult the official model card as DeepSeek fills in documentation. For now, the arrival of a new MIT-licensed flagship gives the open-source community another capable foundation to test, fine-tune, and deploy.

Sources

  • deepseek-ai/DeepSeek-V4-Pro-DSpark

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size842 GB
PrecisionINT8
ArchitectureDeepseekV4ForCausalLM
LicenseMIT
Downloads593.5K
Likes5.6K

Modalities

Text / LLMReasoning
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Text / LLM

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
Atria Dawn Preview
Internlm/Reasoning

InternLM's Atria Dawn Preview Targets Agentic Tasks

A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.

Sep 11, 2026
Unknown/Reasoning

ZGCM-1 arrives as a fully open 7B reasoning model

A compact foundation model targets math reasoning and agentic search with tool use, and its makers are releasing it fully open.

Sep 10, 2026