A User Interface for XTTS-2 Text-Based Voice Cloning using only 10 seconds of speech
☆405Dec 6, 2024Updated last year
Alternatives and similar repositories for xtts2-ui
Users that are interested in xtts2-ui are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Webui for using XTTS and for finetuning it☆893Jan 17, 2025Updated last year
- A simple FastAPI Server to run XTTSv2☆597Jul 21, 2024Updated 2 years ago
- XTTSv2 Extension for oobabooga text-generation-webui☆157Nov 21, 2023Updated 2 years ago
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of adv…☆2,425Jan 9, 2026Updated 7 months ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Slightly improved official version for finetune xtts☆394Apr 3, 2025Updated last year
- A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro,…☆3,235Jul 27, 2026Updated 2 weeks ago
- Running the F5-TTS by ONNX Runtime standalone with GUI☆27Dec 10, 2024Updated last year
- The application allows users to record speech, transcribe it using the Whisper ASR (Automatic Speech Recognition) model, translate the tr…☆15Dec 4, 2023Updated 2 years ago
- A Gradio UI for XTTSv2 and RVC.☆160May 28, 2024Updated 2 years ago
- Pipeline for fast video editing☆13Dec 4, 2023Updated 2 years ago
- A simple extension that uses Bark Text-to-Speech for audio output☆10Nov 20, 2023Updated 2 years ago
- Dia-JAX: A JAX port of Dia, the text-to-speech model for generating realistic dialogue from text with emotion and tone control.☆30May 7, 2025Updated last year
- ☆367Jun 26, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Text-Guided Generation of Full-Body Image with Preserved Reference Face for Customized Animation☆24Jun 24, 2024Updated 2 years ago
- Revolutionize Your Voice with AI Voice Cloner! Transform Your Speech into Your Favorite Celebrity's or Your Customized Voice. Our Cutting…☆74Mar 21, 2025Updated last year
- ☆778Jun 9, 2025Updated last year
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆45,876Aug 16, 2024Updated last year
- Slightly improved official version for finetune xtts☆70Sep 22, 2024Updated last year
- Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with C…☆725Jun 17, 2025Updated last year
- Foundational model for human-like, expressive TTS☆4,204Jul 30, 2024Updated 2 years ago
- Cog wrapper for Coqui / xtts-v2☆84Nov 25, 2024Updated last year
- Real-time Img2img translation! (TouchDesigner+T2Iadapter\_canny+SDXL+turbo\_LoRA)☆20Jan 5, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Instant voice cloning by MIT and MyShell. Audio foundation model.☆37,120Apr 19, 2025Updated last year
- VoxNovel: generate audiobooks giving each character a different voice actor.☆376Jun 8, 2025Updated last year
- Efficient visual programming for AI language models☆356May 13, 2025Updated last year
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 2 months ago
- A simple, high-quality voice conversion tool focused on ease of use and performance.☆3,572Updated this week
- The best looking and most functional webui for RVC related tasks. See website for UI demo:☆227Apr 27, 2024Updated 2 years ago
- Realtime diffusion (LCM-LoRA) from screen capture or webcam, for architecture, using torch and Pyside6☆29Feb 13, 2024Updated 2 years ago
- Installable package for rvc voice inferencing☆10Aug 11, 2024Updated 2 years ago
- replace any object you want on the image with whatever you want☆14Feb 6, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 🔊 Text-Prompted Generative Audio Model with Gradio☆688Nov 23, 2023Updated 2 years ago
- A refactor PersonaPlex to add managed Personalities☆16Apr 5, 2026Updated 4 months ago
- One-shot face animation using webcam, capable of running in real time.☆41May 23, 2024Updated 2 years ago
- A modern web interface for managing and executing ComfyUI workflows with real-time progress tracking and a RESTful API.☆44Jun 4, 2025Updated last year
- Zero-Shot Speech Editing and Text-to-Speech in the Wild☆8,567May 30, 2026Updated 2 months ago
- A webui for different audio related Neural Networks☆1,245May 19, 2025Updated last year
- 🔊 Text-prompted Generative Audio Model☆236Apr 27, 2023Updated 3 years ago