Fine-tuning toolkit for Chatterbox TTS & Chatterbox TURBO models. Supports 23 languages with smart vocabulary extension. Features offline preprocessing, automatic VAD trimming, and voice cloning capabilities. Train custom TTS models with your own dataset in LJSpeech and file-based format.
☆108Jun 1, 2026Updated 2 months ago
Alternatives and similar repositories for chatterbox-finetuning
Users that are interested in chatterbox-finetuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SoTA open-source TTS☆136Jun 7, 2025Updated last year
- With this tool you can create custom TTS dataset from video or audio.☆16Jun 7, 2025Updated last year
- Streaming and Fine-tuning for Chatterbox TTS☆293Jun 15, 2025Updated last year
- ☆27Nov 3, 2025Updated 9 months ago
- SoTA open-source TTS☆26Jul 8, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SoTA open-source TTS☆165Dec 16, 2025Updated 7 months ago
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆20Feb 24, 2026Updated 5 months ago
- High quality text-to-speech based on StyleTTS 2.☆78Apr 6, 2026Updated 3 months ago
- Data preparation utility for the finetuning of OpenAI's Whisper model.☆16Jun 18, 2026Updated last month
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion☆43Sep 9, 2025Updated 10 months ago
- VLLM Port of the Chatterbox TTS model☆379Oct 18, 2025Updated 9 months ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.☆131Jul 25, 2025Updated last year
- FastAPI Implementation of Orpheus TTS streaming Chatbot☆30Jun 19, 2025Updated last year
- Tokenizer for Text to Speech (TTS) models☆14Jan 16, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆254Jan 13, 2026Updated 6 months ago
- SoTA open-source TTS☆23Jun 17, 2025Updated last year
- Real-time voice conversation system with Sesame CSM, featuring web-based audio visualization and GPU acceleration. Educational implementa…☆17Mar 18, 2025Updated last year
- Semantic router for MCP ecosystems - discover and execute tools across multiple servers with intelligent context management.☆22Dec 8, 2025Updated 7 months ago
- VALL-E 2 reproduction☆135Jul 14, 2024Updated 2 years ago
- ☆19Nov 18, 2025Updated 8 months ago
- ☆40Apr 15, 2024Updated 2 years ago
- X-E-Speech: Joint Training Framework of Non-Autoregressive Cross-lingual Emotional Text-to-Speech and Voice Conversion☆112Apr 1, 2024Updated 2 years ago
- This is an unofficial implementation of universal melgan according to https://arxiv.org/abs/2011.09631☆23Aug 15, 2022Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Pure-PyTorch inference for CohereLabs/cohere-transcribe-03-2026 (2B Conformer + Transformer ASR, 14 languages).☆41Apr 29, 2026Updated 3 months ago
- A Weakly Supervised Forced Alignment for disluent speech☆15Nov 12, 2023Updated 2 years ago
- ☆51Apr 20, 2026Updated 3 months ago
- All in one Gradio interface for chatterbox☆20May 31, 2025Updated last year
- ☆205Dec 9, 2024Updated last year
- Redesign of the new version of the QB-Inventory☆13Feb 19, 2024Updated 2 years ago
- ☆14Aug 19, 2024Updated last year
- Aranizer: A Custom Tokenizer based on SentencePiece and BPE tailored for Arabic Language Modeling☆21Aug 4, 2024Updated last year
- Variable Bitrate Residual Vector Quantization for Audio Coding☆54May 1, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- AI SUGGEST is a powerful command-line assistant that leverages AI to provide accurate Linux commands based on natural language queries. S…☆11Aug 22, 2024Updated last year
- Easy fine-tuning for Qwen3-TTS: Fast voice cloning and high-quality multilingual speech synthesis.☆115May 29, 2026Updated 2 months ago
- Unofficial WIP LoRa Finetuning repository for VibeVoice☆372Sep 24, 2025Updated 10 months ago
- [ICASSP 2026]Official code for "Prosody-Guided Harmonic Attention for Phase-Coherent Neural Vocoding in the Complex Spectrum"☆27Jan 22, 2026Updated 6 months ago
- One-command fine-tuning for Qwen3-TTS text-to-speech model with custom voice samples☆29Jan 28, 2026Updated 6 months ago
- Docker XPRA HTML5 Image with OpenGL support for NVIDIA cards☆13Oct 28, 2020Updated 5 years ago
- VyvoTTS: LLM-Based Text-to-Speech Training Framework☆258Apr 8, 2026Updated 3 months ago