The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDeepSeekV4-Flash-0731
DeepSeekText / LLM

DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents

The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.

Jul 31, 2026
NotableMIT
DeepSeek-V4-Flash-0731

DeepSeek has released DeepSeek-V4-Flash-0731, a 304-billion-parameter mixture-of-experts (MoE) model that extends the company's V4 family. According to the model card on Hugging Face, this checkpoint focuses on substantially enhanced agentic capabilities — the ability to plan, call tools, and carry out multi-step tasks rather than simply answering single prompts.

The model is positioned as a text-and-reasoning system, and like DeepSeek's other releases it ships under the permissive MIT license, meaning developers can use, modify, and deploy it commercially with minimal restrictions. That licensing posture has been a consistent differentiator for DeepSeek against more guarded open-weight peers.

Why it matters

The MoE design is central to the appeal here. By routing each token to a subset of specialized experts, a 304B-parameter model can deliver the capacity of a very large network while activating only a fraction of those weights per forward pass — a practical route to strong performance without proportional inference costs.

  • Agent-first tuning: the update targets tool use and multi-step task execution
  • 304B MoE: large total capacity with sparse activation
  • MIT license: unusually open terms for a model of this scale

DeepSeek has not published a context-length figure or detailed benchmark suite alongside this checkpoint, so teams evaluating it for production agent pipelines will want to run their own tests. Still, the "Flash" branding and agentic emphasis signal a model built for responsive, action-oriented deployments rather than pure chat.

Sources

    Get the model

    Hugging Face

    Specs

    Parameters304B · MoE
    Size305.8 GB
    PrecisionINT8
    ArchitectureDeepseekV4ForCausalLM
    LicenseMIT
    Downloads236.1K
    Likes2K

    Modalities

    Text / LLMReasoning
    2 versions — view changelog

    0 comments

    No comments yet. Be the first to weigh in.

    More in Text / LLM

    LongCat-Flash-Lite-Sparse
    Meituan/Text / LLM

    Meituan Ships a Lighter, Sparser LongCat-Flash

    The food-delivery giant's newest open model trims its mixture-of-experts design for more efficient inference under an MIT license.

    Jul 31, 2026
    DeepSeek-V4-Flash-0731
    DeepSeek/Text / LLM

    DeepSeek Refreshes V4-Flash With New 0731 Checkpoint

    The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.

    Jul 31, 2026
    K-EXAONE 2.0 750B-A37B
    LGAI EXAONE/Text / LLM

    LG AI Research debuts K-EXAONE 2.0, a 750B MoE model

    The new mixture-of-experts model activates 37B parameters per token and targets English, Korean, and Spanish reasoning tasks.

    Jul 29, 2026