The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDeepSeekV4-Flash
DeepSeekText / LLM

DeepSeek Releases V4-Flash for Low-Latency Inference

A lighter, faster member of DeepSeek's V4 line arrives on Hugging Face under a permissive MIT license.

Jun 27, 2026
NotableMIT
DeepSeek-V4-Flash

DeepSeek has published DeepSeek-V4-Flash, a text model positioned as the speed-focused option in its V4 family. According to the model's Hugging Face repository, it uses a mixture-of-experts (MoE) design and is aimed at lower-latency inference — the kind of workload where response time matters as much as raw capability.

The release ships under the MIT license, one of the most permissive options available. That means developers can use, modify, and deploy the model commercially with minimal restrictions, continuing DeepSeek's pattern of open, business-friendly licensing that has made its models popular with startups and researchers alike.

Why it matters

Flash-style variants trade some headroom for responsiveness, and they tend to slot into real-world applications more easily than their heavier siblings. A few reasons this release is worth noting:

  • Latency-first design suits chat, agents, and interactive tools where users feel every extra second.
  • MoE architecture can deliver strong throughput by activating only part of the network per token.
  • MIT licensing lowers the legal friction for teams shipping to production.

DeepSeek has not published detailed specifications such as parameter counts or context length alongside this listing, so precise capability claims will have to wait for further documentation or independent testing. For now, the appeal is straightforward: a smaller, quicker V4 that teams can pull down and run without licensing headaches.

As with earlier DeepSeek releases, the practical test will come from the community — how V4-Flash performs on real latency budgets, and how it stacks up against other open lightweight models. Those results should surface quickly once developers start putting it through its paces via the official repository.

Sources

  • deepseek-ai/DeepSeek-V4-Flash-DSpark

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size157.6 GB
PrecisionINT8
ArchitectureDeepseekV4ForCausalLM
LicenseMIT
Downloads2.7M
Likes2K

Modalities

Text / LLM
2 versions — view changelog

0 comments

No comments yet. Be the first to weigh in.

More in Text / LLM

LongCat-Flash-Lite-Sparse
Meituan/Text / LLM

Meituan Ships a Lighter, Sparser LongCat-Flash

The food-delivery giant's newest open model trims its mixture-of-experts design for more efficient inference under an MIT license.

Jul 31, 2026
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Refreshes V4-Flash With New 0731 Checkpoint

The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.

Jul 31, 2026
DeepSeek-V4-Flash-0731
DeepSeek/Text / LLM

DeepSeek Ships V4-Flash, a 304B MoE Tuned for Agents

The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.

Jul 31, 2026