A User Interface for XTTS-2 Text-Based Voice Cloning using only 10 seconds of speech
☆404Dec 6, 2024Updated last year
Alternatives and similar repositories for xtts2-ui
Users that are interested in xtts2-ui are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Webui for using XTTS and for finetuning it☆890Jan 17, 2025Updated last year
- A simple FastAPI Server to run XTTSv2☆595Jul 21, 2024Updated 2 years ago
- XTTSv2 Extension for oobabooga text-generation-webui☆157Nov 21, 2023Updated 2 years ago
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of adv…☆2,416Jan 9, 2026Updated 6 months ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Slightly improved official version for finetune xtts☆393Apr 3, 2025Updated last year
- A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro,…☆3,206Jul 6, 2026Updated 2 weeks ago
- Running the F5-TTS by ONNX Runtime standalone with GUI☆27Dec 10, 2024Updated last year
- The application allows users to record speech, transcribe it using the Whisper ASR (Automatic Speech Recognition) model, translate the tr…☆15Dec 4, 2023Updated 2 years ago
- A Gradio UI for XTTSv2 and RVC.☆159May 28, 2024Updated 2 years ago
- Pipeline for fast video editing☆13Dec 4, 2023Updated 2 years ago
- A simple extension that uses Bark Text-to-Speech for audio output☆10Nov 20, 2023Updated 2 years ago
- Dia-JAX: A JAX port of Dia, the text-to-speech model for generating realistic dialogue from text with emotion and tone control.☆30May 7, 2025Updated last year
- Text-Guided Generation of Full-Body Image with Preserved Reference Face for Customized Animation☆24Jun 24, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Easy to install image colorization using onnx converted DDColor models☆15Apr 2, 2024Updated 2 years ago
- Revolutionize Your Voice with AI Voice Cloner! Transform Your Speech into Your Favorite Celebrity's or Your Customized Voice. Our Cutting…☆74Mar 21, 2025Updated last year
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆45,783Aug 16, 2024Updated last year
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models☆6,314Aug 10, 2024Updated last year
- Slightly improved official version for finetune xtts☆70Sep 22, 2024Updated last year
- Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with C…☆726Jun 17, 2025Updated last year
- Foundational model for human-like, expressive TTS☆4,203Jul 30, 2024Updated last year
- Real-time Img2img translation! (TouchDesigner+T2Iadapter\_canny+SDXL+turbo\_LoRA)☆20Jan 5, 2024Updated 2 years ago
- Cog wrapper for Coqui / xtts-v2☆82Nov 25, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Instant voice cloning by MIT and MyShell. Audio foundation model.☆36,984Apr 19, 2025Updated last year
- Efficient visual programming for AI language models☆357May 13, 2025Updated last year
- VoxNovel: generate audiobooks giving each character a different voice actor.☆371Jun 8, 2025Updated last year
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 2 months ago
- The best looking and most functional webui for RVC related tasks. See website for UI demo:☆223Apr 27, 2024Updated 2 years ago
- A simple, high-quality voice conversion tool focused on ease of use and performance.☆3,501Updated this week
- Realtime diffusion (LCM-LoRA) from screen capture or webcam, for architecture, using torch and Pyside6☆29Feb 13, 2024Updated 2 years ago
- Installable package for rvc voice inferencing☆10Aug 11, 2024Updated last year
- replace any object you want on the image with whatever you want☆14Feb 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆24Jan 22, 2025Updated last year
- A refactor PersonaPlex to add managed Personalities☆16Apr 5, 2026Updated 3 months ago
- Convert Files / Folders / GitHub Repos Into AI / LLM-ready Files☆163Jan 31, 2025Updated last year
- Zero-Shot Speech Editing and Text-to-Speech in the Wild☆8,496May 30, 2026Updated last month
- A modern web interface for managing and executing ComfyUI workflows with real-time progress tracking and a RESTful API.☆43Jun 4, 2025Updated last year
- A webui for different audio related Neural Networks☆1,243May 19, 2025Updated last year
- 🔊 Text-prompted Generative Audio Model☆236Apr 27, 2023Updated 3 years ago