A Gradio-based web UI for voice cloning and voice design, powered by Qwen3-TTS & VibeVoice. Can use Whisper or VibeVoice-ASR for automatic transcription.
☆652May 18, 2026Updated 3 months ago
Alternatives and similar repositories for Voice-Clone-Studio
Users that are interested in Voice-Clone-Studio are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Convert PDFs, EPUBs, DOCX, DOC, and TXT files into high-quality audiobooks using **Qwen3 TTS Voice Model** - an open-source voice synthes…☆1,070Apr 7, 2026Updated 4 months ago
- Qwen3-TTS text-to-speech nodes for ComfyUI with voice cloning, voice design, and fine-tuning UI☆145Apr 25, 2026Updated 3 months ago
- A ComfyUI custom node suite for Qwen3-TTS, supporting 1.7B and 0.6B models, Custom Voice, Voice Design, Voice Cloning and Fine-Tuning.☆297Feb 9, 2026Updated 6 months ago
- A lightweight ComfyUI custom node pack for Qwen3-ASR, providing simple speech‑to‑text workflows with local model caching and optional tim…☆67Jan 31, 2026Updated 6 months ago
- ComfyUI custom nodes for speech, voice cloning, and voice design based on Qwen3-TTS models☆248Jan 30, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Prompt Manager for ComfyUI, with integration with llama.cpp for prompt generation. Allowing users to generate and save prompts, as well a…☆142Updated this week
- Unofficial WIP LoRa Finetuning repository for VibeVoice☆373Sep 24, 2025Updated 10 months ago
- A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.☆5,226Jun 5, 2026Updated 2 months ago
- Easy to use GUI for Qwen TTS 3 for voice creating and cloning☆32Jan 30, 2026Updated 6 months ago
- A Simple Implementation of Qwen3-TTS's ComfyUI☆1,868Jun 3, 2026Updated 2 months ago
- A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwe…☆1,166Updated this week
- ComfyUI custom nodes for Qwen3-ASR (Automatic Speech Recognition) - audio-to-text transcription supporting 52 languages and dialects.☆196Jan 29, 2026Updated 6 months ago
- A professional-grade interface for Qwen3-TTS, designed to unlock the model's full potential with fine-grained control and intuitive workf…☆289Mar 30, 2026Updated 4 months ago
- ☆162Jun 14, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Just a series of workflows I☆34Sep 10, 2025Updated 11 months ago
- ☆35Jan 25, 2026Updated 6 months ago
- ComfyUI node for AudioSR - Versatile Audio Super Resolution upscales audio to 48kHz using latent diffusion☆100Feb 12, 2026Updated 6 months ago
- Custom ComfyUI nodes for downloading, converting, and previewing audio/video from YouTube and 1,000+ other platforms☆32Sep 6, 2025Updated 11 months ago
- A comprehensive WebUI Toolkit for Resemble-AI's Chatterbox☆28Jun 7, 2025Updated last year
- Real-time text-to-speech with Qwen3-TTS☆1,315Jul 17, 2026Updated last month
- Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streamin…☆13,007Mar 17, 2026Updated 5 months ago
- A comprehensive ComfyUI integration for Microsoft's VibeVoice text-to-speech model, enabling high-quality single and multi-speaker voice …☆1,544Feb 18, 2026Updated 6 months ago
- A local-first workspace that keeps your generated media connected to the prompts, models and workflows that created them. Search, compare…☆310Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Graviton: Daisy-Chain ComfyUI workflows. Distribute among multiple GPUs☆44Mar 2, 2026Updated 5 months ago
- On-device AI browser automation using WebLLM. No cloud, no API keys, fully private.☆298Jan 22, 2026Updated 6 months ago
- Automated video dataset creator for Windows using WhisperX and Qwen2-VL☆20Jul 30, 2026Updated 2 weeks ago
- ♡☆23Updated this week
- Port of resemble-ai's DramaBox for ComfyUI☆46May 20, 2026Updated 2 months ago
- OmniVoice TTS nodes for ComfyUI - Zero-shot multilingual text-to-speech with voice cloning, voice design, and multi-speaker dialogue☆535Jun 11, 2026Updated 2 months ago
- MirrorMetrics: How to evaluate Stable Diffusion LoRAs. A visual diagnostic tool to detect overfitting, check dataset quality, and fix tra…☆62Feb 21, 2026Updated 5 months ago
- A TTS that fits in your CPU (and pocket)☆8,705Updated this week
- ComfyUI custom node for the VibeVoice TTS. Expressive, long-form, multi-speaker conversational audio☆595Sep 25, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- TTS model capable of streaming conversational audio in realtime.☆1,162Nov 29, 2025Updated 8 months ago
- FastAPI wrapper around original Vibevoice 1.5B and 7B models, with support for AWQ4 quant☆33Jun 22, 2026Updated last month
- Local-first AI image generation toolkit. Pull models, train LoRAs, generate images. One CLI, no glue code.☆26Jul 25, 2026Updated 3 weeks ago
- Workflows built for my patreon page!☆89Jun 25, 2026Updated last month
- High-Quality Voice Cloning TTS for 600+ Languages☆9,220Updated this week
- Echo-TTS inference codebase☆216Dec 5, 2025Updated 8 months ago
- Zero-shot expressive voice cloning and speech generation. Generate anything from short clips to full-length audiobooks with realistic emo…☆549Jul 7, 2026Updated last month