Fine-tuning toolkit for Chatterbox TTS & Chatterbox TURBO models. Supports 23 languages with smart vocabulary extension. Features offline preprocessing, automatic VAD trimming, and voice cloning capabilities. Train custom TTS models with your own dataset in LJSpeech and file-based format.
☆109Jun 1, 2026Updated 2 months ago
Alternatives and similar repositories for chatterbox-finetuning
Users that are interested in chatterbox-finetuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SoTA open-source TTS☆136Jun 7, 2025Updated last year
- With this tool you can create custom TTS dataset from video or audio.☆16Jun 7, 2025Updated last year
- Streaming and Fine-tuning for Chatterbox TTS☆292Jun 15, 2025Updated last year
- ☆26Nov 3, 2025Updated 9 months ago
- SoTA open-source TTS☆26Jul 8, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- SoTA open-source TTS☆167Dec 16, 2025Updated 8 months ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆20Feb 24, 2026Updated 6 months ago
- High quality text-to-speech based on StyleTTS 2.☆79Apr 6, 2026Updated 4 months ago
- Data preparation utility for the finetuning of OpenAI's Whisper model.☆17Jun 18, 2026Updated 2 months ago
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion☆43Sep 9, 2025Updated 11 months ago
- VLLM Port of the Chatterbox TTS model☆381Oct 18, 2025Updated 10 months ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.☆130Jul 25, 2025Updated last year
- FastAPI Implementation of Orpheus TTS streaming Chatbot☆30Jun 19, 2025Updated last year
- Tokenizer for Text to Speech (TTS) models☆15Jan 16, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆254Jan 13, 2026Updated 7 months ago
- SoTA open-source TTS☆23Jun 17, 2025Updated last year
- Real-time voice conversation system with Sesame CSM, featuring web-based audio visualization and GPU acceleration. Educational implementa…☆17Mar 18, 2025Updated last year
- Semantic router for MCP ecosystems - discover and execute tools across multiple servers with intelligent context management.☆22Dec 8, 2025Updated 8 months ago
- VALL-E 2 reproduction☆135Jul 14, 2024Updated 2 years ago
- ☆19Nov 18, 2025Updated 9 months ago
- ☆40Apr 15, 2024Updated 2 years ago
- X-E-Speech: Joint Training Framework of Non-Autoregressive Cross-lingual Emotional Text-to-Speech and Voice Conversion☆113Apr 1, 2024Updated 2 years ago
- This is an unofficial implementation of universal melgan according to https://arxiv.org/abs/2011.09631☆23Aug 15, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pure-PyTorch inference for CohereLabs/cohere-transcribe-03-2026 (2B Conformer + Transformer ASR, 14 languages).☆42Apr 29, 2026Updated 3 months ago
- A Weakly Supervised Forced Alignment for disluent speech☆15Nov 12, 2023Updated 2 years ago
- ☆52Apr 20, 2026Updated 4 months ago
- All in one Gradio interface for chatterbox☆20May 31, 2025Updated last year
- ☆205Dec 9, 2024Updated last year
- ☆14Aug 19, 2024Updated 2 years ago
- Aranizer: A Custom Tokenizer based on SentencePiece and BPE tailored for Arabic Language Modeling☆21Aug 4, 2024Updated 2 years ago
- Variable Bitrate Residual Vector Quantization for Audio Coding☆55May 1, 2025Updated last year
- AI SUGGEST is a powerful command-line assistant that leverages AI to provide accurate Linux commands based on natural language queries. S…☆11Aug 22, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.☆120May 29, 2026Updated 2 months ago
- Unofficial WIP LoRa Finetuning repository for VibeVoice☆373Sep 24, 2025Updated 10 months ago
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 7 months ago
- One-command fine-tuning for Qwen3-TTS text-to-speech model with custom voice samples☆32Jan 28, 2026Updated 6 months ago
- VyvoTTS: LLM-Based Text-to-Speech Training Framework☆261Aug 9, 2026Updated 2 weeks ago
- KATube is a tool to automate the process of creating datasets for training Text-To-Speech (TTS) and Speech-To-Text (STT) models. From a l…☆26Jul 27, 2024Updated 2 years ago
- superfast text to speech in any voice☆63Feb 16, 2026Updated 6 months ago