Step_Audio_EditX:the first open-source LLM-based audio model excelling at expressive and iterative audio editing—encompassing emotion, speaking style, and paralinguistics—alongside robust zero-shot text-to-speech (TTS) capabilities,try it in comfyUI
☆28Nov 15, 2025Updated 8 months ago
Alternatives and similar repositories for ComfyUI_Step_Audio_EditX_SM
Users that are interested in ComfyUI_Step_Audio_EditX_SM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ComfyUI nodes for Step Audio EditX - State-of-the-art zero-shot voice cloning and audio editing with emotion, style, speed control, and m…☆62Dec 4, 2025Updated 7 months ago
- VoxCPM:Tokenizer-Free TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning,you can use this node,easy infer and easy tr…☆28Updated this week
- A sparse attention kernel supporting mix sparse patterns☆27May 20, 2026Updated 2 months ago
- SoulX-Podcast: Towards Realistic Long-form Podcasts with Dialectal and Paralinguistic Diversity☆92Oct 31, 2025Updated 8 months ago
- Use ‘DICE-Talk’ in ComfyUI,which is a method about 'Correlation-Aware Emotional Talking Portrait Generation'.☆25May 7, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This is a VideoAsPrompt ComfyUI plugin☆21Oct 30, 2025Updated 8 months ago
- InteractAvatar is a novel dual-stream DiT framework that enables talking avatars to perform Grounded Human-Object Interaction (GHOI)☆23Jun 3, 2026Updated last month
- This is a ComfyUI plugin for https://huggingface.co/spaces/Soul-AILab/SoulX-Singer☆17Feb 12, 2026Updated 5 months ago
- This is a Qwen-Image-i2L ComfyUI plugin☆80Dec 12, 2025Updated 7 months ago
- Comfy EverAnimate WIP☆17Jun 5, 2026Updated last month
- ☆20Dec 19, 2025Updated 7 months ago
- FlashVSR:Towards Real-Time Diffusion-Based Streaming Video Super-Resolution,you can use it in comfyUI☆377Feb 23, 2026Updated 4 months ago
- LucidFlux: Caption-Free Universal Image Restoration with a Large-Scale Diffusion Transformer,you can use it in ComfyUI☆62May 27, 2026Updated last month
- enseNova-U1: Unifying Multimodal Understanding and Generation with NEO-Unify Architecture☆71May 15, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation,you can try it in ComfyUI☆43Aug 21, 2025Updated 10 months ago
- ComfyUI custom nodes for LongCat-AudioDiT \ Diffusion-based Zero-Shot Text-to-Speech☆131Apr 4, 2026Updated 3 months ago
- This is a ComfyUI custom node used to convert Qwen-Image LoRA files trained on the ModelScope platform to a format that ComfyUI can recog…☆30Aug 9, 2025Updated 11 months ago
- Implementing FlowEdit, maybe other inversion techniques for the Wan video generation model☆41Feb 28, 2025Updated last year
- interpretability work and exploration for krea☆42Jul 12, 2026Updated last week
- ComfyUI custom_node for ByteDance's InfiniteYou☆11Apr 16, 2025Updated last year
- 不限人数的ComfyUI IndexTTS2节点。ComfyUI IndexTTS2 for multiple participants with no limit on the number of speakers.☆34Sep 14, 2025Updated 10 months ago
- ComfyUI custom nodes for Ovi joint video+audio generation☆47Oct 6, 2025Updated 9 months ago
- ComfyUI node for highly expressive speech and realistic zero-shot voice cloning☆497Apr 24, 2026Updated 2 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Wrapper of DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion, run in diffusers mode☆31Feb 26, 2026Updated 4 months ago
- ☆260Jan 4, 2026Updated 6 months ago
- OmniSVG: A Unified Scalable Vector Graphics Generation Model,you can try it in ComfyUI☆29Dec 5, 2025Updated 7 months ago
- ☆58Dec 25, 2025Updated 6 months ago
- forked from https://github.com/fredconex/ComfyUI-SoundFlow☆24Nov 20, 2025Updated 8 months ago
- A custom ComfyUI node for MiniCPM vision-language models, supporting v4, v4.5, and v4 GGUF formats, enabling high-quality image captionin…☆150Aug 28, 2025Updated 10 months ago
- Qwen3-TTS在ComfyUI的实现,强大语音生成能力,为语音克隆、语音设计、超高质量类人语音生成以及基于自然语言的语音控制提供全面支持。☆130Jan 27, 2026Updated 5 months ago
- 一款ComfyUI扩展节点,能够为您的图像添加各种精美的艺术文字效果,支持丰富的文字样式和特效。☆30Mar 21, 2025Updated last year
- ComfyUI自定义节点,集成ByteDance Sa2VA模型,实现智能图像分割和视觉理解。☆29Dec 17, 2025Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A ComfyUI speech recognition plugin based on [Qwen3-ASR](https://github.com/QwenLM/Qwen3-ASR).☆35Feb 6, 2026Updated 5 months ago
- Quality-preserving per-layer conditioning control for Krea 2 (ComfyUI node). Fork of nova452/ComfyUI-ConditioningKrea2Rebalance — RMS ren…☆109Jun 26, 2026Updated 3 weeks ago
- ☆24Sep 20, 2025Updated 10 months ago
- ☆68Dec 12, 2025Updated 7 months ago
- SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation☆39Mar 31, 2026Updated 3 months ago
- 一个高性能的ComfyUI视频&图像放大插件,利用CUDA加速,支持多GPU、混合精度和Tensor Core优化。☆23Oct 26, 2025Updated 8 months ago
- Fun-CineForge: A Unified Dataset Pipeline and Model for Zero-Shot Movie Dubbing in Diverse Cinematic Scenes☆22Apr 2, 2026Updated 3 months ago