A User Interface for XTTS-2 Text-Based Voice Cloning using only 10 seconds of speech
☆402Dec 6, 2024Updated last year
Alternatives and similar repositories for xtts2-ui
Users that are interested in xtts2-ui are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Webui for using XTTS and for finetuning it☆896Jan 17, 2025Updated last year
- A simple FastAPI Server to run XTTSv2☆595Jul 21, 2024Updated 2 years ago
- XTTSv2 Extension for oobabooga text-generation-webui☆156Nov 21, 2023Updated 2 years ago
- AllTalk is based on the Coqui TTS engine, similar to the Coqui_tts extension for Text generation webUI, however supports a variety of adv…☆2,438Jan 9, 2026Updated 9 months ago
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Slightly improved official version for finetune xtts☆396Apr 3, 2025Updated last year
- A single Gradio + React WebUI with extensions for ACE-Step, OmniVoice, Kimi Audio, Piper TTS, GPT-SoVITS, CosyVoice, XTTSv2, DIA, Kokoro,…☆3,288Sep 7, 2026Updated last month
- Running the F5-TTS by ONNX Runtime standalone with GUI☆29Dec 10, 2024Updated last year
- The application allows users to record speech, transcribe it using the Whisper ASR (Automatic Speech Recognition) model, translate the tr…☆16Sep 21, 2026Updated 2 weeks ago
- A Gradio UI for XTTSv2 and RVC.☆160May 28, 2024Updated 2 years ago
- Pipeline for fast video editing☆13Dec 4, 2023Updated 2 years ago
- A simple extension that uses Bark Text-to-Speech for audio output☆10Nov 20, 2023Updated 2 years ago
- Dia-JAX: A JAX port of Dia, the text-to-speech model for generating realistic dialogue from text with emotion and tone control.☆30May 7, 2025Updated last year
- An NLP driven bot for interacting with Clockify timesheet via different channels☆12Dec 29, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆366Jun 26, 2024Updated 2 years ago
- Text-Guided Generation of Full-Body Image with Preserved Reference Face for Customized Animation☆24Jun 24, 2024Updated 2 years ago
- ☆776Jun 9, 2025Updated last year
- 🐸💬 - a deep learning toolkit for Text-to-Speech, battle-tested in research and production☆46,109Aug 16, 2024Updated 2 years ago
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models☆6,365Aug 10, 2024Updated 2 years ago
- Gui for users who use the coqui-TTS vits model.☆15Sep 16, 2022Updated 4 years ago
- Slightly improved official version for finetune xtts☆70Sep 22, 2024Updated 2 years ago
- Local AI talk with a custom voice based on Zephyr 7B model. Uses RealtimeSTT with faster_whisper for transcription and RealtimeTTS with C…☆736Jun 17, 2025Updated last year
- A simple Perplexity Clone made using Flutter + FastAPI + Gemini☆35Dec 29, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Real-time Img2img translation! (TouchDesigner+T2Iadapter\_canny+SDXL+turbo\_LoRA)☆20Jan 5, 2024Updated 2 years ago
- Cog wrapper for Coqui / xtts-v2☆84Nov 25, 2024Updated last year
- Efficient visual programming for AI language models☆357May 13, 2025Updated last year
- VoxNovel: generate audiobooks giving each character a different voice actor.☆377Jun 8, 2025Updated last year
- A Windows tool to query various LLM AIs. Supports branched conversations, history and summaries among others.☆36May 11, 2026Updated 4 months ago
- The best looking and most functional webui for RVC related tasks. See website for UI demo:☆232Apr 27, 2024Updated 2 years ago
- A simple, high-quality voice conversion tool focused on ease of use and performance.☆3,816Updated this week
- Realtime diffusion (LCM-LoRA) from screen capture or webcam, for architecture, using torch and Pyside6☆30Feb 13, 2024Updated 2 years ago
- replace any object you want on the image with whatever you want☆14Feb 6, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 🔊 Text-Prompted Generative Audio Model with Gradio☆683Nov 23, 2023Updated 2 years ago
- ☆24Jan 22, 2025Updated last year
- A refactor PersonaPlex to add managed Personalities☆16Apr 5, 2026Updated 6 months ago
- Convert Files / Folders / GitHub Repos Into AI / LLM-ready Files☆168Aug 30, 2026Updated last month
- One-shot face animation using webcam, capable of running in real time.☆41May 23, 2024Updated 2 years ago
- Zero-Shot Speech Editing and Text-to-Speech in the Wild☆8,571May 30, 2026Updated 4 months ago
- A webui for different audio related Neural Networks☆1,242May 19, 2025Updated last year