The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDeepSeekV3.1-Base
DeepSeekText / LLM

DeepSeek Releases 671B MoE Model Under MIT License

The new DeepSeek-V3.1-Base is a massive 671-billion-parameter Mixture-of-Experts model designed for efficient, large-scale research and development.

Aug 19, 2025
Major releaseMIT
DeepSeek-V3.2

AI research firm DeepSeek has released DeepSeek-V3.1-Base, a powerful new foundation model that significantly expands the top tier of open-source AI. With a total of 671 billion parameters, it is one of the largest and most capable base models made available to the public under a permissive license.

The model's architecture is a Mixture-of-Experts (MoE), a design that allows for massive parameter counts while managing computational costs. Instead of activating all 671 billion parameters for every task, an MoE model intelligently routes inputs to specialized "expert" subnetworks, making training and inference more efficient than a dense model of equivalent size. The official model card also notes the use of FP8 weights, a lower-precision format that further improves performance and reduces memory requirements.

Why it matters

The release of a model of this scale under the highly permissive MIT license is a major contribution to the open-source community. It provides researchers and developers with a powerful, commercially viable foundation for building specialized applications without the restrictive licensing often attached to state-of-the-art models. This gives organizations a new, high-quality starting point for fine-tuning on proprietary data for complex reasoning and generation tasks.

As a "base" model, DeepSeek-V3.1 is not intended for direct use as a chatbot but is instead optimized for further training and adaptation. Developers can access the model and its components directly from its Hugging Face repository. Its release signals a continuing trend of top-tier AI capabilities becoming more accessible, fostering broader innovation in the field.

Sources

  • deepseek-ai/DeepSeek-V3.1-Base

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters671B · MoE
Active params37B active
Size688.6 GB
PrecisionFP8
ArchitectureDeepseekV3ForCausalLM
LicenseMIT
Downloads2.2M
Likes1.5K

Modalities

Text / LLMReasoning
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Text / LLM

Agnes-3.0-Flash
Agnes AI/Vision-Language

Agnes-3.0-Flash arrives as a multimodal reasoning model

The new release pairs vision-language understanding with a hybrid-attention design aimed at long-context reasoning.

Sep 11, 2026
Atria Dawn Preview
Internlm/Reasoning

InternLM's Atria Dawn Preview Targets Agentic Tasks

A new mixture-of-experts model trained on verified tool interactions arrives as an early preview under an MIT license.

Sep 11, 2026
Unknown/Reasoning

ZGCM-1 arrives as a fully open 7B reasoning model

A compact foundation model targets math reasoning and agentic search with tool use, and its makers are releasing it fully open.

Sep 10, 2026