DeepSeek Announces V4.1 Flash, a Cheaper Reasoning Model
The new mixture-of-experts model is billed as more capable than V4 Pro while costing less to run, and ships under an MIT license.
Company
On the horizon
The new mixture-of-experts model is billed as more capable than V4 Pro while costing less to run, and ships under an MIT license.
Releases
An experimental, MIT-licensed vision-language model brings image understanding to DeepSeek's fast V4 Flash architecture.
The Chinese lab pushes a higher-capability checkpoint of its V4 line to Hugging Face under a permissive MIT license.
The company's newest flagship targets reasoning and coding while keeping a permissive open-source license.
The latest checkpoint in DeepSeek's V4 line leans into agentic workflows while keeping the permissive MIT license.
The MIT-licensed mixture-of-experts model returns in an updated build shipping with FP8 weights for cheaper inference.
The new flagship arrives as a mixture-of-experts system with FP8 weights and open reasoning capabilities under a permissive license.
A lighter, faster member of DeepSeek's V4 line arrives on Hugging Face under a permissive MIT license.
The new flagship model combines a Mixture-of-Experts architecture with a permissive MIT license, positioning it for wide commercial adoption.
The new Mixture of Experts model from the Beijing-based AI lab is optimized for fast, efficient conversational AI and carries a fully permissive license.
The new open vision-language model is designed to extract text and understand structure from complex, multilingual documents.
The new Mixture-of-Experts model from DeepSeek AI combines an efficient FP8 architecture with a fully permissive license for commercial use.
The new vision-language model uses a novel context compression technique to efficiently extract text and structure from complex documents.
The new DeepSeek-V3.1-Base is a massive 671-billion-parameter Mixture-of-Experts model designed for efficient, large-scale research and development.