β205Dec 9, 2024Updated last year
Alternatives and similar repositories for XTTSv2-Finetuning-for-New-Languages
Users that are interested in XTTSv2-Finetuning-for-New-Languages are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is an implementation for train hifigan part of XTTSv2 model using Coqui/TTS.β87Nov 12, 2024Updated last year
- πΈ - A general purpose model trainer, as flexible as it getsβ16Apr 10, 2026Updated 3 months ago
- Montreal Forced Aligner for Vietnameseβ15Oct 23, 2023Updated 2 years ago
- ZMM-TTS: Zero-shot Multilingual and Multispeaker Speech Synthesis Conditioned on Self-supervised Discrete Speech Representationsβ183Mar 6, 2024Updated 2 years ago
- finetune llm part for spark-tts modelβ125Mar 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- SoTA open-source TTSβ136Jun 7, 2025Updated last year
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ2,295Jun 10, 2026Updated last month
- Just another FastSpeech 2 but cleaner code :)β29Jun 28, 2024Updated 2 years ago
- KABooks is a tool to automate the process of creating datasets for training Text-To-Speech (TTS) and Speech-To-Text (STT) models. Using aβ¦β13Mar 24, 2023Updated 3 years ago
- An unofficial PyTorch implementation of VALL-Eβ88Aug 3, 2025Updated 11 months ago
- A transcribed speech dataset in Wolof, Pulaar and Sereer, to support agriculture. Funded by Lacuna Fund.β20Mar 26, 2026Updated 3 months ago
- Forced alignment decoder for Whisper.β16Mar 13, 2024Updated 2 years ago
- LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesisβ660Jan 21, 2026Updated 6 months ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.β131Jul 25, 2025Updated 11 months ago
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Hosts text-to-speech corpus and speech synthesizers for African languages.β19May 31, 2023Updated 3 years ago
- A Vietnamese handwriting recognition projectβ16Feb 21, 2024Updated 2 years ago
- [TAFFC 2025] The official implementation of EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vecβ¦β129Updated this week
- [Computer Speech & Language] A transformer-based spelling error correction framework for Bangla and resource scarce Indic languagesβ14Aug 9, 2024Updated last year
- β31Oct 29, 2024Updated last year
- Scripts to create speech corpora from open.bibleβ13Jan 3, 2022Updated 4 years ago
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ22Apr 7, 2024Updated 2 years ago
- VietTTS: An Open-Source Vietnamese Text to Speechβ88Dec 23, 2025Updated 6 months ago
- Finetune VITS and MMS using HuggingFace's toolsβ202Mar 31, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Text to speech alignment using CTC forced alignmentβ523Jul 12, 2026Updated last week
- Webui for using XTTS and for finetuning itβ890Jan 17, 2025Updated last year
- HiFTNet wav/audio super-resolution 16/24 kHz to 48 kHzβ24Jan 2, 2024Updated 2 years ago
- Vi_G2P or ViG2P: G2P package for Vietnamese: based on vPhon and phonology knowledge to convert Raw text - Graphoneme to IPAβ109Jun 21, 2024Updated 2 years ago
- Slightly improved official version for finetune xttsβ393Apr 3, 2025Updated last year
- β161Apr 23, 2025Updated last year
- β45Sep 19, 2024Updated last year
- π Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. π§π₯π Advanced audio processing.β262Jun 10, 2024Updated 2 years ago
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversionβ43Sep 9, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- unofficial vits2-TTS implementation in pytorchβ548Mar 28, 2024Updated 2 years ago
- A neural speech codec based on discrete WavLM representationsβ26Aug 28, 2024Updated last year
- VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Designβ642Sep 11, 2023Updated 2 years ago
- VyvoTTS: LLM-Based Text-to-Speech Training Frameworkβ257Apr 8, 2026Updated 3 months ago
- Official implementation of the TTS model Lina-Speechβ178Jan 9, 2025Updated last year
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matchingβ1,016Dec 2, 2025Updated 7 months ago
- β21Mar 7, 2023Updated 3 years ago