ACE-Step: A Step Towards Music Generation Foundation Model
☆250May 28, 2025Updated last year
Alternatives and similar repositories for ComfyUI_ACE-Step
Users that are interested in ComfyUI_ACE-Step are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A ComfyUI node containing multiple audio processing tools.☆103Jul 7, 2025Updated last year
- Generative Motion Latent Flow Matching for Audio-driven Talking Portrait☆271Jun 5, 2026Updated 2 months ago
- Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation. A node for ComfyUI.☆154May 30, 2025Updated last year
- SongGeneration:High-Quality Song Generation with Multi-Preference Alignment (SOTA),you can try VRAM>12G☆161Mar 21, 2026Updated 5 months ago
- Lightweight and Efficient, 🎧Ultra High-Quality Voice Cloning, Chinese and English.☆211Jun 11, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple Characters,try it in comfyUI ,if your VRAM >24G☆80Jun 24, 2025Updated last year
- A ComfyUI extention for BAGEL(Unified Model for Multimodal Understanding and Generation)☆188Jul 31, 2026Updated last month
- ACE-Step: A Step Towards Music Generation Foundation Model☆4,794Feb 15, 2026Updated 6 months ago
- YuE is a groundbreaking series of open-source foundation models designed for music generation, specifically for transforming lyrics into …☆190Feb 24, 2025Updated last year
- An node for ComfyUI that implements AceStep 1.5 SFT (Supervised Fine-Tuning), a high-quality music generation model. This node replicates…☆56May 11, 2026Updated 3 months ago
- ComfyUI-Bagel is now available in ComfyUI, BAGEL is an open‑source multimodal foundation model with 7B active parameters (14B total) trai…☆30May 28, 2025Updated last year
- A Text To Speech node using Step-Audio-TTS in ComfyUI. Can speak, rap, sing, or clone voice.☆167May 23, 2025Updated last year
- Seed-VC voice or sing conversion.☆67Jun 11, 2025Updated last year
- ☆227Feb 5, 2026Updated 6 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- DiffuEraser is a diffusion model for video Inpainting, you can use it in ComfyUI☆279Mar 27, 2026Updated 5 months ago
- ☆575Feb 1, 2026Updated 6 months ago
- DreamO native implementation for ComfyUI☆182Aug 22, 2025Updated last year
- Prompt Generator for Video, Audio, Image, and Text. A node for ComfyUI. Including Deepseek, Alibaba Cloud Qwen, Google Gemini, and locall…☆56Jul 11, 2025Updated last year
- 使用IndexTTS模型在ComfyUI中实现高质量文本到语音转换的自定义节点。支持中文和英文文本,可以基于参考音频复刻声音特征。☆748Aug 15, 2026Updated 2 weeks ago
- ComfyUI Custom Node for HeartMuLa AI Music Generation and Transcript Text☆434Mar 23, 2026Updated 5 months ago
- ComfyUI Plugin of Nunchaku☆2,920Feb 19, 2026Updated 6 months ago
- Sonic is a method about ' Shifting Focus to Global Audio Perception in Portrait Animation',you can use it in comfyUI☆1,141May 4, 2026Updated 3 months ago
- Everything-Reactivity in ComfyUI (audio, MIDI, motion, proximity, and more). Animate and manipulate images, masks, videos, audio, and mo…☆859Mar 20, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Creates prompts for Video Models by sequence analysis and prompting using Qwen2.5-VL models from Alibaba.☆57Apr 2, 2025Updated last year
- ☆149Dec 14, 2025Updated 8 months ago
- A custom ComfyUI node for MiniCPM vision-language models, supporting v4, v4.5, and v4 GGUF formats, enabling high-quality image captionin…☆154Aug 28, 2025Updated last year
- Mel-Band RoFormer for Music Source Separation☆248Jan 30, 2026Updated 7 months ago
- Portrait Tools: Facial detection cropping, alignment, ID photo, etc☆21Jun 15, 2025Updated last year
- ☆75Apr 30, 2025Updated last year
- This is a ComfyUI plug-in for TencentARC/IC-Custom☆36Sep 3, 2025Updated 11 months ago
- A ComfyUI integration for FireRedTTS‑2, a real-time multi-speaker TTS system enabling high-quality, emotionally expressive dialogue and m…☆43Sep 16, 2025Updated 11 months ago
- The nodes detached from [ComfyUI Layer Style](https://github.com/chflame163/ComfyUI_LayerStyle) are mainly those with complex requirement…☆707Apr 8, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- for tile the image for advanced control or modification☆1,033Jul 7, 2026Updated last month
- ComfyUI nodes to crop before sampling and stitch back after sampling that speed up inpainting☆1,147Aug 21, 2026Updated last week
- ☆369Jan 27, 2026Updated 7 months ago
- The core component of Stand-In, the preprocessor, is essential—only images processed through it can fully unlock the capabilities of Stan…☆157Aug 21, 2025Updated last year
- ComfyUI Wrapper for HiDream☆478Apr 22, 2025Updated last year
- IndexTTS Voice Cloning: Supports two-person dialogue☆553Nov 7, 2025Updated 9 months ago
- You can using StoryDiffusion in ComfyUI☆515Oct 11, 2025Updated 10 months ago