super expressive prompting model based on ltx2.3
☆467May 23, 2026Updated last month
Alternatives and similar repositories for DramaBox
Users that are interested in DramaBox are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Zero-shot expressive voice cloning and speech generation. Generate anything from short clips to full-length audiobooks with realistic emo…☆534Jul 7, 2026Updated 2 weeks ago
- Port of resemble-ai's DramaBox for ComfyUI☆43May 20, 2026Updated 2 months ago
- comfyui node for dramabox tts☆19May 25, 2026Updated last month
- LTX Director and a variety of other custom ComfyUI nodes and workflows☆1,789Jul 15, 2026Updated last week
- KVAE-Audio: a continuous full-band audio waveform autoencoder☆98Jun 30, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- High-Quality Voice Cloning TTS for 600+ Languages☆8,371Updated this week
- ☆605Updated this week
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆346Jun 24, 2026Updated 3 weeks ago
- MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fi…☆3,853Jun 22, 2026Updated 3 weeks ago
- A high quality and fast TTS repository☆517Dec 22, 2025Updated 6 months ago
- Scaled diffusion transformer for text-to-speech synthesis (DiT + T5Gemma2 conditioning, TorchTitan & Megatron backends, tested up to 1024…☆24Mar 29, 2026Updated 3 months ago
- Echo-TTS inference codebase☆204Dec 5, 2025Updated 7 months ago
- Soprano-Factory: Train your own 2000x realtime text-to-speech model☆249Jan 13, 2026Updated 6 months ago
- ☆918Jul 10, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆552Apr 3, 2026Updated 3 months ago
- Real-time stream editing pipeline powered by the FLUX.2-klein-4B model, optimized for consumer GPUs☆420Jun 13, 2026Updated last month
- ☆365Aug 28, 2025Updated 10 months ago
- X-Voice☆176Jun 5, 2026Updated last month
- Soprano: Instant, Ultra-Realistic Text-to-Speech☆1,250Jan 15, 2026Updated 6 months ago
- An inference-time, plug-and-play method for temporal control in multi-event generation☆184Apr 26, 2026Updated 2 months ago
- Text-to-audio and video-to-audio using Sony AI's Woosh foundation model.☆127May 7, 2026Updated 2 months ago
- ☆302Jul 22, 2025Updated 11 months ago
- The inference and trainging code for WordVoice.☆48Updated this week
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Fast audio super resolution from 16khz to 48khz.☆215Jan 3, 2026Updated 6 months ago
- Unofficial WIP LoRa Finetuning repository for VibeVoice☆367Sep 24, 2025Updated 9 months ago
- A powerful 3B-parameter, LLM-based Reinforcement Learning audio edit model excels at editing emotion, speaking style, and paralinguistics…☆951Apr 9, 2026Updated 3 months ago
- Custom ComfyUI nodes for film grain, color matching, and video enhancement.☆535Updated this week
- ☆376Jul 13, 2026Updated last week
- Open Source Speech Language Model☆1,007May 11, 2026Updated 2 months ago
- VLLM Port of the Chatterbox TTS model☆379Oct 18, 2025Updated 9 months ago
- [ArXiv 26] EverAnimate: Minute-Scale Human Animation via Latent Flow Restoration☆72Jun 4, 2026Updated last month
- A ComfyUI custom node integration for local multi-engine multi-language Text-to-Speech and Voice Conversion. Supports: RVC, Echo-TTS, Qwe…☆1,103Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for "OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation"☆149Jun 18, 2026Updated last month
- ☆461Nov 2, 2025Updated 8 months ago
- Inference server for MioTTS, a lightweight and fast LLM-based TTS model.☆197Feb 14, 2026Updated 5 months ago
- LLaSA: Scaling Train-time and Inference-time Compute for LLaMA-based Speech Synthesis☆660Jan 21, 2026Updated 6 months ago
- An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System☆145Nov 15, 2025Updated 8 months ago
- GLM-TTS: Controllable & Emotion-Expressive Zero-shot TTS with Multi-Reward Reinforcement Learning☆1,043Apr 10, 2026Updated 3 months ago
- ☆90Dec 31, 2025Updated 6 months ago