π π€ Pip installable package for StyleTTS 2 human-level text-to-speech and voice cloning
β159Jul 15, 2024Updated 2 years ago
Alternatives and similar repositories for StyleTTS2
Users that are interested in StyleTTS2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Modelsβ6,329Aug 10, 2024Updated 2 years ago
- β99Apr 27, 2024Updated 2 years ago
- Fine Tune the Style-TTS2 Voice Modelβ267Jun 17, 2025Updated last year
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversionβ43Sep 9, 2025Updated 11 months ago
- π Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. π§π₯π Advanced audio processing.β262Jun 10, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Application of MB-iSTFT-VITS components to vits2_pytorchβ135Dec 29, 2025Updated 7 months ago
- Create Unmute voice embeddingsβ26Nov 15, 2025Updated 8 months ago
- πΈ - A general purpose model trainer, as flexible as it getsβ16Apr 10, 2026Updated 4 months ago
- Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictionsβ270Jan 13, 2025Updated last year
- Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorchβ517Dec 20, 2025Updated 7 months ago
- Official Implementation of StyleTTSβ468Jan 13, 2025Updated last year
- Controllable and fast Text-to-Speech for over 7000 languages!β2,211Jan 25, 2026Updated 6 months ago
- β14Aug 19, 2024Updated last year
- An Open Source text-to-speech system built by inverting Whisper.β4,629Dec 14, 2025Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Vocoder NSF-HiFiGAN (Moved into deepaudio)β56Dec 11, 2022Updated 3 years ago
- Quick hack job to allow use with Sillytavern. This works for me, some further updates are expected to expose more settings to sillytavernβ11May 30, 2024Updated 2 years ago
- An unofficial PyTorch implementation of VALL-Eβ88Aug 3, 2025Updated last year
- VITS-based zero-shot TTS system varying with diverse style/speaker conditioning methods.β36Sep 21, 2022Updated 3 years ago
- Text-Guided Generation of Full-Body Image with Preserved Reference Face for Customized Animationβ24Jun 24, 2024Updated 2 years ago
- β12Mar 28, 2024Updated 2 years ago
- A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro,β¦β3,236Jul 27, 2026Updated 2 weeks ago
- This is the official code for ACM CIKM 2025 Paper: ParaStyleTTS: Toward Efficient and Robust Paralinguistic Style Control for Expressive β¦β59Dec 21, 2025Updated 7 months ago
- β12Mar 18, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available β’ AdRun AI, ML, and HPC workloads on powerful cloud GPUsβwithout limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- β31Oct 29, 2024Updated last year
- This project includes a Python script for fine-tuning a text-to-speech (TTS) model. The script utilizes custom datasets and use CUDA for β¦β14Oct 4, 2024Updated last year
- β16Apr 23, 2024Updated 2 years ago
- A ggml (C++) re-implementation of tortoise-ttsβ193Aug 20, 2024Updated last year
- Try to replicate the architecture of MiniMaxTTS mentioned in it's technical reportβ47Sep 2, 2025Updated 11 months ago
- A very simple implementation of edge_tts w/ RVC for oobabooga text-generation-webui.β42Jan 26, 2024Updated 2 years ago
- This project is based on SadTalker to implement video lip synthesis.β14Jan 9, 2024Updated 2 years ago
- Oobabooga extension for Bark TTSβ118Nov 23, 2023Updated 2 years ago
- DEX-TTS: Diffusion-based EXpressive TTS with Style Modeling on Time Variabilityβ108Jan 17, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- β19May 2, 2024Updated 2 years ago
- Incorporating AutoVocoder to MB-iSTFT-VITSβ47Dec 1, 2022Updated 3 years ago
- Orpheus TTS Server with streaming support (TTFB ~160ms)β26Sep 21, 2025Updated 10 months ago
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of advβ¦β2,427Jan 9, 2026Updated 7 months ago
- Putting flows on top of neural transducers for better TTSβ64Updated this week
- ComfyUI style LDM patching in A1111β52Jun 11, 2024Updated 2 years ago
- a Frontier Japanese Speech Generation netβ65May 15, 2025Updated last year