β205Dec 9, 2024Updated last year
Alternatives and similar repositories for XTTSv2-Finetuning-for-New-Languages
Users that are interested in XTTSv2-Finetuning-for-New-Languages are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is an implementation for train hifigan part of XTTSv2 model using Coqui/TTS.β87Nov 12, 2024Updated last year
- πΈ - A general purpose model trainer, as flexible as it getsβ16Apr 10, 2026Updated 5 months ago
- Montreal Forced Aligner for Vietnameseβ15Oct 23, 2023Updated 2 years ago
- ZMM-TTS: Zero-shot Multilingual and Multispeaker Speech Synthesis Conditioned on Self-supervised Discrete Speech Representationsβ186Mar 6, 2024Updated 2 years ago
- finetune llm part for spark-tts modelβ126Mar 25, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ2,322Jun 10, 2026Updated 3 months ago
- SoTA open-source TTSβ137Jun 7, 2025Updated last year
- Just another FastSpeech 2 but cleaner code :)β29Jun 28, 2024Updated 2 years ago
- KABooks is a tool to automate the process of creating datasets for training Text-To-Speech (TTS) and Speech-To-Text (STT) models. Using aβ¦β13Mar 24, 2023Updated 3 years ago
- An unofficial PyTorch implementation of VALL-Eβ88Aug 3, 2025Updated last year
- A transcribed speech dataset in Wolof, Pulaar and Sereer, to support agriculture. Funded by Lacuna Fund.β22Mar 26, 2026Updated 5 months ago
- Forced alignment decoder for Whisper.β16Mar 13, 2024Updated 2 years ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.β130Jul 25, 2025Updated last year
- LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesisβ658Jan 21, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Hosts text-to-speech corpus and speech synthesizers for African languages.β21May 31, 2023Updated 3 years ago
- A Vietnamese handwriting recognition projectβ17Feb 21, 2024Updated 2 years ago
- [TAFFC 2025] The official implementation of EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vecβ¦β131Jul 16, 2026Updated 2 months ago
- [Computer Speech & Language] A transformer-based spelling error correction framework for Bangla and resource scarce Indic languagesβ14Sep 12, 2026Updated last week
- β31Oct 29, 2024Updated last year
- Scripts to create speech corpora from open.bibleβ13Jan 3, 2022Updated 4 years ago
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ22Apr 7, 2024Updated 2 years ago
- VietTTS: An Open-Source Vietnamese Text to Speechβ86Dec 23, 2025Updated 8 months ago
- Text to speech alignment using CTC forced alignmentβ563Sep 7, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Finetune VITS and MMS using HuggingFace's toolsβ205Mar 31, 2024Updated 2 years ago
- Webui for using XTTS and for finetuning itβ897Jan 17, 2025Updated last year
- HiFTNet wav/audio super-resolution 16/24 kHz to 48 kHzβ24Jan 2, 2024Updated 2 years ago
- Vi_G2P or ViG2P: G2P package for Vietnamese: based on vPhon and phonology knowledge to convert Raw text - Graphoneme to IPAβ109Jun 21, 2024Updated 2 years ago
- Slightly improved official version for finetune xttsβ395Apr 3, 2025Updated last year
- β172Apr 23, 2025Updated last year
- β45Sep 19, 2024Updated 2 years ago
- π Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. π§π₯π Advanced audio processing.β263Jun 10, 2024Updated 2 years ago
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversionβ43Sep 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- unofficial vits2-TTS implementation in pytorchβ549Mar 28, 2024Updated 2 years ago
- VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Designβ647Sep 11, 2023Updated 3 years ago
- VyvoTTS: LLM-Based Text-to-Speech Training Frameworkβ260Updated this week
- Official implementation of the TTS model Lina-Speechβ177Jan 9, 2025Updated last year
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matchingβ1,069Dec 2, 2025Updated 9 months ago
- β21Mar 7, 2023Updated 3 years ago
- Speech synthesis (TTS) in low-resource languages by training from scratch with Fastpitch and fine-tuning with HifiGanβ70Dec 5, 2023Updated 2 years ago