β204Dec 9, 2024Updated last year
Alternatives and similar repositories for XTTSv2-Finetuning-for-New-Languages
Users that are interested in XTTSv2-Finetuning-for-New-Languages are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is an implementation for train hifigan part of XTTSv2 model using Coqui/TTS.β87Nov 12, 2024Updated last year
- πΈ - A general purpose model trainer, as flexible as it getsβ16Oct 2, 2026Updated last week
- Montreal Forced Aligner for Vietnameseβ15Oct 23, 2023Updated 2 years ago
- ZMM-TTS: Zero-shot Multilingual and Multispeaker Speech Synthesis Conditioned on Self-supervised Discrete Speech Representationsβ186Mar 6, 2024Updated 2 years ago
- finetune llm part for spark-tts modelβ126Mar 25, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform β’ AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ2,335Oct 2, 2026Updated last week
- SoTA open-source TTSβ139Jun 7, 2025Updated last year
- Just another FastSpeech 2 but cleaner code :)β29Jun 28, 2024Updated 2 years ago
- KABooks is a tool to automate the process of creating datasets for training Text-To-Speech (TTS) and Speech-To-Text (STT) models. Using aβ¦β13Mar 24, 2023Updated 3 years ago
- An unofficial PyTorch implementation of VALL-Eβ88Aug 3, 2025Updated last year
- A transcribed speech dataset in Wolof, Pulaar and Sereer, to support agriculture. Funded by Lacuna Fund.β22Mar 26, 2026Updated 6 months ago
- Forced alignment decoder for Whisper.β16Mar 13, 2024Updated 2 years ago
- A TTS model capable of generating ultra-realistic dialogue in one pass.β130Jul 25, 2025Updated last year
- LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesisβ658Jan 21, 2026Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Hosts text-to-speech corpus and speech synthesizers for African languages.β21May 31, 2023Updated 3 years ago
- [TAFFC 2025] The official implementation of EmoSphere++: Emotion-Controllable Zero-Shot Text-to-Speech via Emotion-Adaptive Spherical Vecβ¦β132Jul 16, 2026Updated 2 months ago
- [Computer Speech & Language] A transformer-based spelling error correction framework for Bangla and resource scarce Indic languagesβ14Sep 12, 2026Updated 3 weeks ago
- β31Oct 29, 2024Updated last year
- Scripts to create speech corpora from open.bibleβ13Jan 3, 2022Updated 4 years ago
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ22Apr 7, 2024Updated 2 years ago
- VietTTS: An Open-Source Vietnamese Text to Speechβ88Dec 23, 2025Updated 9 months ago
- Text to speech alignment using CTC forced alignmentβ566Sep 7, 2026Updated last month
- Finetune VITS and MMS using HuggingFace's toolsβ205Mar 31, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Webui for using XTTS and for finetuning itβ896Jan 17, 2025Updated last year
- HiFTNet wav/audio super-resolution 16/24 kHz to 48 kHzβ24Jan 2, 2024Updated 2 years ago
- Vi_G2P or ViG2P: G2P package for Vietnamese: based on vPhon and phonology knowledge to convert Raw text - Graphoneme to IPAβ110Jun 21, 2024Updated 2 years ago
- Slightly improved official version for finetune xttsβ396Apr 3, 2025Updated last year
- β174Apr 23, 2025Updated last year
- β45Sep 19, 2024Updated 2 years ago
- π Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. π§π₯π Advanced audio processing.β263Jun 10, 2024Updated 2 years ago
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversionβ43Sep 9, 2025Updated last year
- unofficial vits2-TTS implementation in pytorchβ548Mar 28, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture Designβ646Sep 11, 2023Updated 3 years ago
- VyvoTTS: LLM-Based Text-to-Speech Training Frameworkβ263Sep 17, 2026Updated 3 weeks ago
- Official implementation of the TTS model Lina-Speechβ177Jan 9, 2025Updated last year
- Fast and High-Quality Zero-Shot Text-to-Speech with Flow Matchingβ1,082Dec 2, 2025Updated 10 months ago
- β21Mar 7, 2023Updated 3 years ago
- Speech synthesis (TTS) in low-resource languages by training from scratch with Fastpitch and fine-tuning with HifiGanβ70Dec 5, 2023Updated 2 years ago
- [ACMMM'2024] Generative Expressive Conversational Speech Synthesisβ45Oct 28, 2024Updated last year