Resemble AI Releases Chatterbox Turbo for Open TTS
The new text-to-speech model focuses on performance and offers voice cloning capabilities for English under a permissive MIT license.
Resemble AI, a company specializing in synthetic voice technology, has released a new open-source model named Chatterbox Turbo. The model is designed for high-performance text-to-speech (TTS) generation in English, targeting developers who need fast and efficient voice output in their applications.
Beyond standard speech synthesis, Chatterbox Turbo includes voice cloning capabilities, allowing users to create speech in a specific target voice from a short audio sample. The entire project is available on Hugging Face under the permissive MIT license, encouraging wide use and modification in both academic and commercial projects.
This release adds another strong contender to the rapidly growing field of open-source speech generation. While proprietary APIs have dominated the high-quality TTS space, permissively licensed models like Chatterbox Turbo provide a crucial, self-hostable alternative for developers. This move from a commercial provider signals a broader trend of companies contributing foundational models back to the community.
Developers interested in experimenting with the model can find the necessary code and instructions on the official Hugging Face repository.
Sources
- Visit
ResembleAI/chatterbox-turbo
Hugging Face
More in Text → Speech
StepFun's StepAudio 3 Realtime targets live voice AI
The audio-language foundation model builds a listen-converse-think-act loop aimed at natural, low-latency spoken interaction.
StepFun's StepAudio 3 Gen Unifies TTS and Music
A single discrete autoregressive model handles speech, voice design, sound effects, and music generation.
Breeze-TTS-2 Brings Open Voice Cloning to English
BreezeBlue's second-generation text-to-speech model pairs voice cloning with controllable direction, all under an open release on Hugging Face.
0 comments
No comments yet. Be the first to weigh in.