πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
β16Sep 19, 2026Updated last week
Alternatives and similar repositories for YourTTS
Users that are interested in YourTTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The implementation of paper "SpeechTripleNet: End-to-End Disentangled Speech Representation Learning for Content, Timbre and Prosody"β33Nov 23, 2023Updated 2 years ago
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ36Mar 10, 2022Updated 4 years ago
- SC-CNN: Effective Speaker Conditioning Method for Zero-Shot Multi-Speaker Text-to-Speech Systemsβ39Nov 1, 2023Updated 2 years ago
- β12Jun 14, 2024Updated 2 years ago
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervβ¦β41Jan 6, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- GitHub repository for AudioToolAgentβ21Feb 13, 2026Updated 7 months ago
- PitchVC: Pitch Conditioned Any-to-Many Voice Conversionβ35Jun 6, 2024Updated 2 years ago
- Implementation of SoundtStream from the paper: "SoundStream: An End-to-End Neural Audio Codec"β13Jan 27, 2025Updated last year
- Onnx compatible styletts2 codeβ16Apr 4, 2026Updated 5 months ago
- Inference codebase for "Cacophony: An Improved Contrastive Audio-Text Model". Preprint: https://arxiv.org/abs/2402.06986β49Jan 19, 2026Updated 8 months ago
- [TOMM 2024] Automatic Lyric Transcription and Automatic Music Transcription from Multimodal Singingβ28Aug 30, 2024Updated 2 years ago
- β10Mar 22, 2023Updated 3 years ago
- Implementation of Acoustic BPE (Shen et al., 2024), extended for RVQ-based Neural Audio Codecsβ77Dec 3, 2025Updated 9 months ago
- DUSTED: Spoken-Term Discovery using Discrete Speech Unitsβ17Oct 2, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The open source code for SimpleSpeech seriesβ147Oct 8, 2024Updated last year
- Official Pytorch Implementation of "Diff-HierVC: Diffusion-based Hierarchical Voice Conversion with Robust Pitch Generation and Masked Prβ¦β238Jul 3, 2024Updated 2 years ago
- used to evaluate wavenet vocoder by rmse f0, MCD, rmse ap...β15Jan 20, 2020Updated 6 years ago
- [Early Alpha] A unified framework for text-to-speech, voice conversion, automatic speech recognition, audio classification, voice activitβ¦β21Jan 10, 2025Updated last year
- ε¦δΉ η¨PyTorchεδ½εθ―β17Mar 17, 2019Updated 7 years ago
- Megatts2 use HierSpeechpp's vocoderβ17Dec 2, 2024Updated last year
- β19Mar 2, 2024Updated 2 years ago
- Zalo AI Challenge 2020 - Top 2 @ Voice Verificationβ15Oct 4, 2022Updated 3 years ago
- Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024β70Feb 19, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MaskGCT demo pageβ14Feb 9, 2025Updated last year
- Official Repository for ICASSP 2024 Paper "SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription"β35Dec 6, 2024Updated last year
- BUPT Software Engineering Homeworkβ16Jun 14, 2020Updated 6 years ago
- A real-time voice conversion model based on VITS.β16Aug 1, 2024Updated 2 years ago
- TriAAN-VC: Triple Adaptive Attention Normalization for Any-to-Any Voice Conversionβ146Jan 15, 2024Updated 2 years ago
- E2E TTS using Conditional Flow Matching (Experimental*)β71Nov 10, 2023Updated 2 years ago
- VITS-based zero-shot TTS system varying with diverse style/speaker conditioning methods.β36Sep 21, 2022Updated 4 years ago
- Unofficial Implementation of "Liu, W., Li, A., Wang, X., Yuan, M., Chen, Y., Zheng, C., & Li, X. (2022). A Neural Beamspace-Domain Filterβ¦β19Oct 21, 2022Updated 3 years ago
- Scripts to work with an intracranial EEG dataset of speech production.β29Jan 12, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β38Aug 24, 2026Updated last month
- NLP project, based on Bert/UNILMβ23Dec 29, 2019Updated 6 years ago
- dog-can-sing-songβ159Jul 24, 2026Updated 2 months ago
- QuickVC: Any-to-many Voice Conversion Using Inverse Short-time Fourier Transform for Faster Conversionβ261Jul 13, 2023Updated 3 years ago
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speakeβ¦β57Aug 7, 2023Updated 3 years ago
- Repo for the IDESSAI 2024 course on modeling audio with discrete tokens.β13Sep 13, 2024Updated 2 years ago
- β14Jun 16, 2023Updated 3 years ago