Voice Recognition

Chatterbox – family of text-to-speech models

Chatterbox is a family of text-to-speech models from Resemble AI. It supports speech generation, voice cloning, and multilingual synthesis, with models targeting different performance and hardware requirements.

The family includes Chatterbox Multilingual V3 for speech synthesis across more than 20 languages, Chatterbox-Turbo for lower-latency English generation, and the smaller Chatterbox-Nano for resource-constrained and CPU-based deployments. Turbo and Nano also understand paralinguistic tags such as laughs, coughs, and chuckles.

Generated audio incorporates Resemble AI’s PerTh neural watermarking technology, which is designed to remain detectable after common audio processing and compression.

This is free and open source software.

Key Features

  • Generate natural-sounding speech from text.
  • Clone voices from reference audio and preserve speaker characteristics across languages.
  • Supports multilingual speech synthesis across more than 20 languages.
  • Offers Turbo and Nano models for lower-latency and resource-constrained deployments.
  • Supports paralinguistic tags and embeds neural watermarks in generated audio.

Website: github.com/resemble-ai/chatterbox
Support:
Developer: Resemble AI
License: MIT License

Chatterbox is written in Python. Learn Python with our recommended free books and free tutorials.


Related Software

Speech Tools
PiperFast, local neural text to speech system
TortoiseMulti-voice text-to-speech system trained with an emphasis on quality
Coqui TTSOffers pretrained models in more than 1,100 different languages
BarkTransformer-based text-to-audio model.
Dia1.6B parameter text to speech model
FestivalGeneral multi-lingual speech synthesis system
PraatSpeechAnalyserSoftware for speech analysis and synthesis
Speech NoteSpeech to Text, Text to Speech and Machine Translation
Mimic 3Lightweight Text to Speech engine
OrcaScreenReaderScriptable screen reader
MeloTTSHigh-quality multi-lingual text-to-speech library
Parler-TTSLightweight text-to-speech (TTS) model
FliteSmall, fast run time text to speech synthesis engine
RHVoiceGives the visually impaired a synthesis voice with their screen reader
eSpeak NGContinuation of the eSpeak project
eSpeakSpeech synthesizer using a formant synthesis method
Orpheus-TTS-FastAPIHigh-performance self-hosted text-to-speech server
GespeakerGTK-based frontend for eSpeak
VoiceGenSimple text-to-speech application
GlateGoogle Translator and Text To Speech Service

Read our verdict in the software roundup.


Best Free and Open Source Software Explore our comprehensive directory of recommended free and open source software. Our carefully curated collection spans every major software category.

This directory is part of our ongoing series of informative articles for Linux enthusiasts. It features hundreds of detailed reviews, along with open source alternatives to proprietary solutions from major corporations such as Google, Microsoft, Apple, Adobe, IBM, Cisco, Oracle, and Autodesk.

You’ll also find interesting projects to try, hardware coverage, free programming books and tutorials, and much more.

Discovered a useful open source Linux program that we haven’t covered yet? Let us know by completing this form.
Subscribe

Please read our Comment Policy before commenting.

Notify of
guest
0 Comments
Oldest
Newest Most Voted