The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestMistral AIMini-4B-Realtime-2602
Mistral AISpeech → Text

Mistral Enters Speech AI with Voxtral Mini Model

The company, known for its powerful text models, has released its first open-source speech recognition system designed for real-time, multilingual transcription.

Jan 21, 2026
NotableApache 2.0
Voxtral Mini 4B Realtime

Mistral AI, a company that has rapidly built a reputation for its powerful open-source text models, has released Voxtral Mini 4B Realtime, its first publicly available model for automatic speech recognition (ASR).

This new 4-billion-parameter system is designed specifically for real-time, multilingual speech-to-text applications. Its focus on low-latency performance makes it a candidate for tasks like live captioning, meeting transcription, and voice-activated assistants where immediate feedback is critical.

A New Modality for Mistral

The release signals a significant expansion for the Paris-based AI lab. While previously focused exclusively on text generation with models like Mistral 7B and Mixtral, the company is now entering the competitive audio AI space. This move positions Voxtral as an open-source alternative to established ASR systems, including OpenAI's popular Whisper model.

By releasing Voxtral Mini under a permissive Apache 2.0 license, Mistral continues its strategy of providing foundational tools for developers. The model is now available for download and experimentation on the Hugging Face Hub, allowing the community to build upon and integrate it into new voice-powered applications.

Sources

  • mistralai/Voxtral-Mini-4B-Realtime-2602

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters4B
Languages13 languages
Size8.9 GB
PrecisionBF16
ArchitectureVoxtralRealtimeForConditionalGeneration
LicenseAPACHE-2.0
Downloads2M
Likes985

Modalities

Speech → Text

0 comments

No comments yet. Be the first to weigh in.

More in Speech → Text

StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
VibeVoice-ASR-Streaming-7B
Microsoft/Speech → Text

Microsoft's VibeVoice ASR brings streaming speech-to-text

A 7-billion-parameter model targets real-time, multilingual transcription with an open release on Hugging Face.

Sep 2, 2026
s1-mini
Superwhisper/Speech → Text

Superwhisper's s1-mini polishes raw speech-to-text output

A compact Qwen3-based model tackles the unglamorous cleanup work that makes transcripts readable.

Aug 12, 2026