🐍 🤖 Pip installable package for StyleTTS 2 human-level text-to-speech and voice cloning
☆159Jul 15, 2024Updated 2 years ago
Alternatives and similar repositories for StyleTTS2
Users that are interested in StyleTTS2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models☆6,344Aug 10, 2024Updated 2 years ago
- ☆99Apr 27, 2024Updated 2 years ago
- Fine Tune the Style-TTS2 Voice Model☆267Jun 17, 2025Updated last year
- [EMNLP 2025 Findings] Official code for EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion☆43Sep 9, 2025Updated 11 months ago
- 🔊 Create labeled datasets, enhance audio quality, identify speakers, support diverse dataset types. 🎧👥📊 Advanced audio processing.☆262Jun 10, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Application of MB-iSTFT-VITS components to vits2_pytorch☆135Dec 29, 2025Updated 8 months ago
- Create Unmute voice embeddings☆26Nov 15, 2025Updated 9 months ago
- 🐸 - A general purpose model trainer, as flexible as it gets☆16Apr 10, 2026Updated 4 months ago
- Phoneme-Level BERT for Enhanced Prosody of Text-to-Speech with Grapheme Predictions☆269Jan 13, 2025Updated last year
- Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in Pytorch☆517Dec 20, 2025Updated 8 months ago
- Official Implementation of StyleTTS☆469Jan 13, 2025Updated last year
- Controllable and fast Text-to-Speech for over 7000 languages!☆2,207Jan 25, 2026Updated 7 months ago
- ☆14Aug 19, 2024Updated 2 years ago
- An Open Source text-to-speech system built by inverting Whisper.☆4,646Dec 14, 2025Updated 8 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Vocoder NSF-HiFiGAN (Moved into deepaudio)☆56Dec 11, 2022Updated 3 years ago
- Quick hack job to allow use with Sillytavern. This works for me, some further updates are expected to expose more settings to sillytavern☆11May 30, 2024Updated 2 years ago
- An unofficial PyTorch implementation of VALL-E☆88Aug 3, 2025Updated last year
- VITS-based zero-shot TTS system varying with diverse style/speaker conditioning methods.☆36Sep 21, 2022Updated 3 years ago
- Text-Guided Generation of Full-Body Image with Preserved Reference Face for Customized Animation☆24Jun 24, 2024Updated 2 years ago
- ☆13Mar 28, 2024Updated 2 years ago
- A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro,…☆3,251Updated this week
- Text to speech Plugin for Flow☆14Aug 26, 2025Updated last year
- This is the official code for ACM CIKM 2025 Paper: ParaStyleTTS: Toward Efficient and Robust Paralinguistic Style Control for Expressive …☆60Dec 21, 2025Updated 8 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Mar 18, 2024Updated 2 years ago
- ☆31Oct 29, 2024Updated last year
- This project includes a Python script for fine-tuning a text-to-speech (TTS) model. The script utilizes custom datasets and use CUDA for …☆14Oct 4, 2024Updated last year
- ☆16Apr 23, 2024Updated 2 years ago
- A ggml (C++) re-implementation of tortoise-tts☆192Aug 20, 2024Updated 2 years ago
- Try to replicate the architecture of MiniMaxTTS mentioned in it's technical report☆47Sep 2, 2025Updated last year
- A very simple implementation of edge_tts w/ RVC for oobabooga text-generation-webui.☆42Jan 26, 2024Updated 2 years ago
- This project is based on SadTalker to implement video lip synthesis.☆14Jan 9, 2024Updated 2 years ago
- Oobabooga extension for Bark TTS☆118Nov 23, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- DEX-TTS: Diffusion-based EXpressive TTS with Style Modeling on Time Variability☆108Jan 17, 2025Updated last year
- ☆18May 2, 2024Updated 2 years ago
- Incorporating AutoVocoder to MB-iSTFT-VITS☆47Dec 1, 2022Updated 3 years ago
- Orpheus TTS Server with streaming support (TTFB ~160ms)☆26Sep 21, 2025Updated 11 months ago
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of adv…☆2,431Jan 9, 2026Updated 7 months ago
- Putting flows on top of neural transducers for better TTS☆64Updated this week
- ComfyUI style LDM patching in A1111☆52Jun 11, 2024Updated 2 years ago