The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestNVIDIA1.0
NVIDIAText → Speech

NVIDIA opens Magpie TTS for multilingual voice agents

The company releases open weights for a low-latency, multilingual text-to-speech model aimed at real-time conversational systems.

Aug 10, 2026
NotableOther

NVIDIA has released Magpie TTS Multilingual, an open-weights text-to-speech model designed for building responsive voice agents. According to NVIDIA's announcement on Hugging Face, the model emphasizes low latency and multilingual synthesis — two qualities that matter most when speech is generated on the fly rather than pre-rendered.

The pitch is straightforward: teams building conversational systems often have to choose between the quality of a hosted API and the control of running their own stack. By publishing the weights, NVIDIA is offering the latter, giving developers full deployment control over where and how the model runs.

Why it matters

Voice agents live or die by responsiveness. A few hundred extra milliseconds of delay makes a conversation feel stilted, so a model tuned for real-time generation across multiple languages fills a practical gap for anyone building assistants, IVR replacements, or accessibility tools.

  • Open weights, so the model can be self-hosted rather than accessed only through an API
  • Multilingual output for cross-language deployments
  • Low-latency design aimed at real-time conversational use

NVIDIA distributes the model under a custom license, so teams should review the terms before shipping. Details, including deployment guidance, are laid out in the company's post.

Sources

    Get the model


    Specs

    Languagesmultilingual
    LicenseOTHER

    Modalities

    Text → Speech

    0 comments

    No comments yet. Be the first to weigh in.

    More in Text → Speech

    Unknown/Text → Speech

    Nari Labs Ships Qwen3-Based TTS and ASR Models

    The startup pairs speech synthesis and recognition built on Qwen3, pitching accuracy, low latency, and lower cost.

    Sep 14, 2026
    StepFun/Text → Speech

    StepFun's StepAudio 3 Realtime targets live voice AI

    The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

    Sep 11, 2026
    StepFun/Text → Speech

    StepFun's StepAudio 3 Gen Unifies TTS and Music

    A single discrete autoregressive model handles speech, voice design, sound effects, and music generation.

    Sep 10, 2026