The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestUnknown3
UnknownText → Speech

Nari Labs Ships Qwen3-Based TTS and ASR Models

The startup pairs speech synthesis and recognition built on Qwen3, pitching accuracy, low latency, and lower cost.

Sep 14, 2026
NotableOther

Nari Labs has released a pair of speech models built on Qwen3: a text-to-speech (TTS) system and an automatic speech recognition (ASR) system. The company frames the launch around three practical goals that matter to anyone building voice applications — accuracy, low latency, and cost — rather than raw novelty.

According to Nari Labs, the models lead on Coval voice AI benchmarks, a suite that evaluates conversational voice performance. By offering both synthesis and recognition, Nari is targeting the full round-trip of a voice agent: understanding what a user says and responding in natural speech.

Why it matters

Voice pipelines have historically stitched together separate vendors for ASR and TTS, each adding latency and per-minute costs that compound in real-time settings. A single provider tuning both ends on a common foundation — here, Alibaba's Qwen3 family — can smooth those seams.

  • Accuracy: framed as competitive on Coval benchmarks
  • Latency: a priority for real-time, interactive use
  • Cost: positioned as a differentiator against incumbent APIs

Nari Labs is best known for Dia, its earlier open TTS work, so a Qwen3-based expansion into ASR signals broader ambitions in the voice stack. The release leaves some specifics — parameter counts, licensing terms, and language coverage — worth watching as more details surface.

Sources

  • Show HN: Nari Qwen3-TTS and Qwen3-ASR – High accuracy, low latency and cost

    Hacker News

    Visit

Get the model

Hacker News

Specs

LicenseOTHER

Modalities

Speech → TextText → Speech

0 comments

No comments yet. Be the first to weigh in.

More in Text → Speech

StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
StepFun/Text → Speech

StepFun's StepAudio 3 Gen Unifies TTS and Music

A single discrete autoregressive model handles speech, voice design, sound effects, and music generation.

Sep 10, 2026
Breeze-TTS-2
BreezeBlue/Text → Speech

Breeze-TTS-2 Brings Open Voice Cloning to English

BreezeBlue's second-generation text-to-speech model pairs voice cloning with controllable direction, all under an open release on Hugging Face.

Aug 25, 2026