The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestDeepSeekV4-Flash-Vision-Exp
DeepSeekVision-Language

DeepSeek adds vision to its V4 Flash line

An experimental, MIT-licensed vision-language model brings image understanding to DeepSeek's fast V4 Flash architecture.

Aug 31, 2026
NotableMIT
DeepSeek-V4-Flash-Vision-Exp

DeepSeek has published DeepSeek-V4-Flash-Vision-Exp, an experimental vision-language variant of its V4 Flash model. As the name signals, this is a work-in-progress release rather than a polished flagship, but it extends the company's fast, lightweight line into multimodal territory.

The model handles both image and text inputs, pairing visual understanding with the text generation of the underlying Flash architecture. Like other recent DeepSeek releases, it uses a mixture-of-experts design, which activates only part of the network per token to keep inference costs down while preserving capacity.

Why it matters

DeepSeek has built a reputation for shipping capable models under permissive terms, and this release continues that pattern:

  • It carries an MIT license, allowing broad commercial and research use.
  • It targets the Flash tier, oriented toward efficiency rather than maximum scale.
  • It marks the family's first step into vision-language tasks.

The "Exp" tag is a clear caution: DeepSeek is signaling that behavior, capabilities, and stability may shift before any production-grade version arrives. For developers, that makes this a model to experiment with rather than deploy, but it offers an early look at how DeepSeek intends to fold vision into its faster models. Details on parameter counts, context length, and benchmarks were not specified in the release; those interested should consult the model card directly.

Sources

  • deepseek-ai/DeepSeek-V4-Flash-Vision-Exp

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Size306.7 GB
PrecisionINT8
ArchitectureDeepseekV4ForCausalLM
LicenseMIT
Likes160

Modalities

Text / LLMVision-Language

0 comments

No comments yet. Be the first to weigh in.

More in Vision-Language

GLM-5.3-Flash
Zhipu AI/Text / LLM

Zhipu releases GLM-5.3-Flash under MIT license

A speed-tuned member of the GLM-5.3 family arrives with open weights and mixture-of-experts design aimed at fast, low-cost inference.

Aug 25, 2026
WeMM
Tencent/Embeddings

Tencent's WeMM-Embedding-9B Unifies Text, Image and Video

The WeChat team releases a 9-billion-parameter multimodal embedding model that maps three modalities into one shared vector space.

Aug 25, 2026
Thomson-1.0-Small
Thomsonreuters/Text / LLM

Thomson Reuters enters the model race with Thomson-1.0

The information giant's first frontier model is a mixture-of-experts system tuned on its proprietary legal, tax, and news data.

Aug 18, 2026