Step_Audio_EditX:the first open-source LLM-based audio model excelling at expressive and iterative audio editing—encompassing emotion, speaking style, and paralinguistics—alongside robust zero-shot text-to-speech (TTS) capabilities,try it in comfyUI
☆28Nov 15, 2025Updated 10 months ago
Alternatives and similar repositories for ComfyUI_Step_Audio_EditX_SM
Users that are interested in ComfyUI_Step_Audio_EditX_SM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- VoxCPM:Tokenizer-Free TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning,you can use this node,easy infer and easy tr…☆31Aug 14, 2026Updated last month
- A sparse attention kernel supporting mix sparse patterns☆29May 20, 2026Updated 4 months ago
- SoulX-Podcast: Towards Realistic Long-form Podcasts with Dialectal and Paralinguistic Diversity☆94Oct 31, 2025Updated 11 months ago
- Use ‘DICE-Talk’ in ComfyUI,which is a method about 'Correlation-Aware Emotional Talking Portrait Generation'.☆24May 7, 2025Updated last year
- This is a VideoAsPrompt ComfyUI plugin☆21Oct 30, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- InteractAvatar is a novel dual-stream DiT framework that enables talking avatars to perform Grounded Human-Object Interaction (GHOI)☆25Jun 3, 2026Updated 4 months ago
- This is a ComfyUI plugin for https://huggingface.co/spaces/Soul-AILab/SoulX-Singer☆18Feb 12, 2026Updated 7 months ago
- Comfy EverAnimate WIP☆17Jun 5, 2026Updated 4 months ago
- ☆20Dec 19, 2025Updated 9 months ago
- This is a Qwen-Image-i2L ComfyUI plugin☆81Dec 12, 2025Updated 9 months ago
- FlashVSR:Towards Real-Time Diffusion-Based Streaming Video Super-Resolution,you can use it in comfyUI☆397Feb 23, 2026Updated 7 months ago
- LucidFlux: Caption-Free Universal Image Restoration with a Large-Scale Diffusion Transformer,you can use it in ComfyUI☆63May 27, 2026Updated 4 months ago
- enseNova-U1: Unifying Multimodal Understanding and Generation with NEO-Unify Architecture☆93Sep 28, 2026Updated last week
- StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation,you can try it in ComfyUI☆44Aug 21, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is a ComfyUI custom node used to convert Qwen-Image LoRA files trained on the ModelScope platform to a format that ComfyUI can recog…☆30Aug 9, 2025Updated last year
- Implementing FlowEdit, maybe other inversion techniques for the Wan video generation model☆41Feb 28, 2025Updated last year
- interpretability work and exploration for krea☆60Jul 12, 2026Updated 2 months ago
- 不限人数的ComfyUI IndexTTS2节点。ComfyUI IndexTTS2 for multiple participants with no limit on the number of speakers.☆34Sep 14, 2025Updated last year
- ComfyUI custom nodes for Ovi joint video+audio generation☆46Oct 6, 2025Updated last year
- ComfyUI node for highly expressive speech and realistic zero-shot voice cloning☆514Aug 6, 2026Updated 2 months ago
- Wrapper of DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion, run in diffusers mode☆31Feb 26, 2026Updated 7 months ago
- ☆275Jan 4, 2026Updated 9 months ago
- OmniSVG: A Unified Scalable Vector Graphics Generation Model,you can try it in ComfyUI☆32Dec 5, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆58Dec 25, 2025Updated 9 months ago
- forked from https://github.com/fredconex/ComfyUI-SoundFlow☆28Nov 20, 2025Updated 10 months ago
- A custom ComfyUI node for MiniCPM vision-language models, supporting v4, v4.5, and v4 GGUF formats, enabling high-quality image captionin…☆157Aug 28, 2025Updated last year
- Qwen3-TTS在ComfyUI的实现,强大语音生成能力,为语音克隆、语音设计、超高质量类人语音生成以及基于自然语言的语音控制提供全面支持。☆138Jan 27, 2026Updated 8 months ago
- 一款ComfyUI扩展节点,能够为您的图像添加各种精美的艺术文字效果,支持丰富的文字样式和特效。☆30Aug 19, 2026Updated last month
- ☆24Sep 20, 2025Updated last year
- ComfyUI自定义节点,集成ByteDance Sa2VA模型,实现智能图像分割和视觉理解。☆29Dec 17, 2025Updated 9 months ago
- Quality-preserving per-layer conditioning control for Krea 2 (ComfyUI node). Fork of nova452/ComfyUI-ConditioningKrea2Rebalance — RMS ren…☆166Jun 26, 2026Updated 3 months ago
- A ComfyUI speech recognition plugin based on [Qwen3-ASR](https://github.com/QwenLM/Qwen3-ASR).☆37Feb 6, 2026Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆68Dec 12, 2025Updated 9 months ago
- ☆54Updated this week
- SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation☆40Mar 31, 2026Updated 6 months ago
- Fun-CineForge: A Unified Dataset Pipeline and Model for Zero-Shot Movie Dubbing in Diverse Cinematic Scenes☆20Apr 2, 2026Updated 6 months ago
- ☆46Oct 26, 2024Updated last year
- 一个高性能的ComfyUI视频&图像放大插件,利用CUDA加速,支持多GPU、混合精度和Tensor Core优化。☆23Oct 26, 2025Updated 11 months ago
- A Text To Speech node using Kokoro TTS in ComfyUI. Supports 8 languages and 150 voices☆33Jun 2, 2025Updated last year