The Open Weights
LatestModelsLeaderboardsCompanies
Subscribe
The Open Weights

The daily record of open-source AI. New model releases, leaderboards, and what's coming next — written for people who ship.

Refreshed every 12 hours

Discover

  • Latest releases
  • New today
  • Trending models

Browse

  • All models
  • Companies
  • Categories
  • Leaderboards

About

  • About
  • Editorial policy
  • RSS feed
  • Newsletter

© 2026 The Open Weights. An independent publication.

PrivacyTermsSMSAggregated by Claude · curated by humans.
LatestMaya Research1.0
Maya ResearchText → Speech

Veena TTS Model Targets Indian Languages with Llama Base

Maya Research has released a 3-billion-parameter model designed to generate natural-sounding speech in Hindi and English.

Jun 24, 2025
NotableOther
Veena

Maya Research has introduced Veena, a new open-source model for text-to-speech (TTS) synthesis. The model is specifically designed to address the need for high-quality voice generation in major Indian languages.

What sets Veena apart is its foundation on a Llama-style architecture. The 3-billion-parameter model is trained to generate natural-sounding speech from text in both Hindi and English, catering to the nuances of Indian accents and dialects. This architectural choice leverages the powerful text-processing capabilities of large language models for the distinct task of audio generation.

The release marks a significant step for open-source AI in a region where high-quality, accessible models have been less common. By focusing on widely spoken Indian languages, Veena could enable a new range of applications, from localized voice assistants and accessibility tools to automated content creation for one of the world's largest digital audiences.

The model, its weights, and usage instructions are available for download on the Hugging Face Hub. It is released under a custom license, and potential users should review the terms before implementation.

Sources

  • maya-research/Veena

    Hugging Face

    Visit

Get the model

Hugging Face

Specs

Parameters3B
Sample rate24 kHz
LanguagesHindi, English
Size7.6 GB
PrecisionBF16
ArchitectureLlamaForCausalLM
LicenseOTHER
Downloads15K
Likes240

Modalities

Text → Speech

0 comments

No comments yet. Be the first to weigh in.

More in Text → Speech

StepFun/Text → Speech

StepFun's StepAudio 3 Realtime targets live voice AI

The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.

Sep 11, 2026
StepFun/Text → Speech

StepFun's StepAudio 3 Gen Unifies TTS and Music

A single discrete autoregressive model handles speech, voice design, sound effects, and music generation.

Sep 10, 2026
Breeze-TTS-2
BreezeBlue/Text → Speech

Breeze-TTS-2 Brings Open Voice Cloning to English

BreezeBlue's second-generation text-to-speech model pairs voice cloning with controllable direction, all under an open release on Hugging Face.

Aug 25, 2026