The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestNyralabs2.0-large
NyralabsSpeech → Text

CrisperWhisper 2.0 Large targets verbatim transcription

A Whisper-based ASR model that keeps every filler word and stamps timestamps to the individual word, now covering English and German.

Jul 15, 2026
NotableOther
CrisperWhisper 2.0 Large

Nyra Labs has released CrisperWhisper 2.0 Large, an automatic speech recognition model built on OpenAI's Whisper architecture but tuned for a specific goal: transcribing exactly what was said, including the ums, ahs, and false starts that most systems quietly delete.

The model supports English and German, and its headline features are verbatim output with disfluency handling and precise word-level timestamps. That combination makes it a better fit for use cases where fidelity matters more than tidy prose.

Why it matters

Most commercial and open ASR pipelines are optimized to produce clean, readable text, smoothing over hesitations and repetitions. That is helpful for note-taking but a problem for other work:

  • Linguistics and speech research, where disfluencies are the object of study
  • Legal and medical records that require an accurate account of what was spoken
  • Precise alignment tasks like subtitling and audio editing, which lean on word-level timing

By keeping the disfluencies and pinning timestamps to each word, CrisperWhisper 2.0 fills a niche that general-purpose transcription tools tend to ignore. As with any Whisper derivative, teams will want to validate accuracy on their own audio and confirm the licensing terms on the model page before deploying it in production.

Sources

  • nyralabs/CrisperWhisper2.0_large

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Languagesen, de
Size3.1 GB
PrecisionBF16
ArchitectureWhisperForConditionalGenerationWithAttentionLoss
LicenseOTHER
Downloads15.2K
Likes116

Modalities

Speech → Text

0 comments

No comments yet. Be the first to weigh in.

More in Speech → Text

StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
VibeVoice-ASR-Streaming-7B
Microsoft/Speech → Text

Microsoft's VibeVoice ASR brings streaming speech-to-text

A 7-billion-parameter model targets real-time, multilingual transcription with an open release on Hugging Face.

Sep 2, 2026
s1-mini
Superwhisper/Speech → Text

Superwhisper's s1-mini polishes raw speech-to-text output

A compact Qwen3-based model tackles the unglamorous cleanup work that makes transcripts readable.

Aug 12, 2026