Building with Chatterbox TTS, Voice Cloning & Watermarking
In this video, I look at the new Chatterbox TTS from Resemble.AI and how it's improving open-source text-to-speech with its impressive voice cloning and emotion control capabilities. We explore its features, including zero-shot voice cloning that requires only a few seconds of audio, and its unique ability to adjust the emotional intensity of speech. Colab: https://dripl.ink/Vxs8D Blog: https://www.resemble.ai/chatterbox/ Hugging Face Spaces: https://huggingface.co/spaces/ResembleAI/Chatterbox Hugging Face: https://huggingface.co/ResembleAI/chatterbox GitHub: Chatterbox-TTS-Extended https://github.com/petermg/Chatterbox-TTS-Extended For more tutorials on using LLMs and building agents, check out my Patreon Patreon: https://www.patreon.com/SamWitteveen Twitter: https://x.com/Sam_Witteveen 🕵️ Interested in building LLM Agents? Fill out the form below Building LLM Agents Form: https://drp.li/dIMes 👨💻Github: https://github.com/samwit/llm-tutorials ⏱️Time Stamps: 00:00 Intro 00:24 Resemble.AI - Chatterbox 01:53 Samples 04:53 Hugging Face: Chatterbox 05:22 Demo 06:26 Adding Exaggeration 08:56 Voice Cloning 13:00 Chatterbox TTS Extended Github 14:07 Hugging Face: Chatterbox GGUF