πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and production
β16Feb 4, 2026Updated 5 months ago
Alternatives and similar repositories for YourTTS
Users that are interested in YourTTS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The implementation of paper "SpeechTripleNet: End-to-End Disentangled Speech Representation Learning for Content, Timbre and Prosody"β33Nov 23, 2023Updated 2 years ago
- πΈπ¬ - a deep learning toolkit for Text-to-Speech, battle-tested in research and productionβ37Mar 10, 2022Updated 4 years ago
- SC-CNN: Effective Speaker Conditioning Method for Zero-Shot Multi-Speaker Text-to-Speech Systemsβ39Nov 1, 2023Updated 2 years ago
- β12Jun 14, 2024Updated 2 years ago
- PyTorch implementation of the ICASSP-24 paper: "Improving Audio Captioning Models with Fine-grained Audio Features, Text Embedding Supervβ¦β41Jan 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- GitHub repository for AudioToolAgentβ20Feb 13, 2026Updated 5 months ago
- PitchVC: Pitch Conditioned Any-to-Many Voice Conversionβ35Jun 6, 2024Updated 2 years ago
- Implementation of SoundtStream from the paper: "SoundStream: An End-to-End Neural Audio Codec"β13Jan 27, 2025Updated last year
- Onnx compatible styletts2 codeβ16Apr 4, 2026Updated 3 months ago
- Inference codebase for "Cacophony: An Improved Contrastive Audio-Text Model". Preprint: https://arxiv.org/abs/2402.06986β49Jan 19, 2026Updated 6 months ago
- [TOMM 2024] Automatic Lyric Transcription and Automatic Music Transcription from Multimodal Singingβ28Aug 30, 2024Updated last year
- β10Mar 22, 2023Updated 3 years ago
- DUSTED: Spoken-Term Discovery using Discrete Speech Unitsβ17Oct 2, 2024Updated last year
- Implementation of Acoustic BPE (Shen et al., 2024), extended for RVQ-based Neural Audio Codecsβ76Dec 3, 2025Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The open source code for SimpleSpeech seriesβ147Oct 8, 2024Updated last year
- Official Pytorch Implementation of "Diff-HierVC: Diffusion-based Hierarchical Voice Conversion with Robust Pitch Generation and Masked Prβ¦β237Jul 3, 2024Updated 2 years ago
- used to evaluate wavenet vocoder by rmse f0, MCD, rmse ap...β15Jan 20, 2020Updated 6 years ago
- [Early Alpha] A unified framework for text-to-speech, voice conversion, automatic speech recognition, audio classification, voice activitβ¦β22Jan 10, 2025Updated last year
- ε¦δΉ η¨PyTorchεδ½εθ―β17Mar 17, 2019Updated 7 years ago
- Megatts2 use HierSpeechpp's vocoderβ18Dec 2, 2024Updated last year
- β19Mar 2, 2024Updated 2 years ago
- Zalo AI Challenge 2020 - Top 2 @ Voice Verificationβ15Oct 4, 2022Updated 3 years ago
- Repository for the paper "Combining audio control and style transfer using latent diffusion", accepted at ISMIR 2024β67Feb 19, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MaskGCT demo pageβ14Feb 9, 2025Updated last year
- A real-time voice conversion model based on VITS.β16Aug 1, 2024Updated last year
- BUPT Software Engineering Homeworkβ16Jun 14, 2020Updated 6 years ago
- Official Repository for ICASSP 2024 Paper "SynthTab: Leveraging Synthesized Data for Guitar Tablature Transcription"β33Dec 6, 2024Updated last year
- TriAAN-VC: Triple Adaptive Attention Normalization for Any-to-Any Voice Conversionβ146Jan 15, 2024Updated 2 years ago
- E2E TTS using Conditional Flow Matching (Experimental*)β71Nov 10, 2023Updated 2 years ago
- VITS-based zero-shot TTS system varying with diverse style/speaker conditioning methods.β36Sep 21, 2022Updated 3 years ago
- Unofficial Implementation of "Liu, W., Li, A., Wang, X., Yuan, M., Chen, Y., Zheng, C., & Li, X. (2022). A Neural Beamspace-Domain Filterβ¦β19Oct 21, 2022Updated 3 years ago
- Scripts to work with an intracranial EEG dataset of speech production.β29Jan 12, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- NLP project, based on Bert/UNILMβ23Dec 29, 2019Updated 6 years ago
- dog-can-sing-songβ59Updated this week
- QuickVC: Any-to-many Voice Conversion Using Inverse Short-time Fourier Transform for Faster Conversionβ261Jul 13, 2023Updated 3 years ago
- Unofficial Pytorch implementation of SNAC: Speaker-normalized affine coupling layer in flow-based architecture for zero-shot multi-speakeβ¦β57Aug 7, 2023Updated 2 years ago
- Repo for the IDESSAI 2024 course on modeling audio with discrete tokens.β13Sep 13, 2024Updated last year
- β14Jun 16, 2023Updated 3 years ago
- β10Aug 16, 2024Updated last year