Instructions to use ResembleAI/chatterbox with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Chatterbox
How to use ResembleAI/chatterbox with Chatterbox:
# pip install chatterbox-tts import torchaudio as ta from chatterbox.tts import ChatterboxTTS model = ChatterboxTTS.from_pretrained(device="cuda") text = "Ezreal and Jinx teamed up with Ahri, Yasuo, and Teemo to take down the enemy's Nexus in an epic late-game pentakill." wav = model.generate(text) ta.save("test-1.wav", wav, model.sr) # If you want to synthesize with a different voice, specify the audio prompt AUDIO_PROMPT_PATH="YOUR_FILE.wav" wav = model.generate(text, audio_prompt_path=AUDIO_PROMPT_PATH) ta.save("test-2.wav", wav, model.sr) - Inference
- Notebooks
- Google Colab
- Kaggle
Difference between Chatterbox Multilingual versions
#54
by MariaB99 - opened
Hi,
What is the real difference between versions 2 and 3 of the model? I didn't find the exact documentation about those differences
All files are literally identical. They clai all sorts of improvements in teerms of accuracy w/ the weights themselves, but I don't buy it. This is a tired strategy to stay relevant. They inflate their weights, anuwau. These are realy fp32 weights and a simply model.half(() script cuts the size and ram requirements in half for this model. They do that to make the models difficult to self-host, to drive people to their paid api website. These guys are con artists.