Step_Audio_EditX:the first open-source LLM-based audio model excelling at expressive and iterative audio editing—encompassing emotion, speaking style, and paralinguistics—alongside robust zero-shot text-to-speech (TTS) capabilities,try it in comfyUI
☆28Nov 15, 2025Updated 9 months ago
Alternatives and similar repositories for ComfyUI_Step_Audio_EditX_SM
Users that are interested in ComfyUI_Step_Audio_EditX_SM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ComfyUI nodes for Step Audio EditX - State-of-the-art zero-shot voice cloning and audio editing with emotion, style, speed control, and m…☆63Dec 4, 2025Updated 8 months ago
- VoxCPM:Tokenizer-Free TTS for Context-Aware Speech Generation and True-to-Life Voice Cloning,you can use this node,easy infer and easy tr…☆29Aug 14, 2026Updated 2 weeks ago
- A sparse attention kernel supporting mix sparse patterns☆28May 20, 2026Updated 3 months ago
- SoulX-Podcast: Towards Realistic Long-form Podcasts with Dialectal and Paralinguistic Diversity☆93Oct 31, 2025Updated 9 months ago
- Use ‘DICE-Talk’ in ComfyUI,which is a method about 'Correlation-Aware Emotional Talking Portrait Generation'.☆24May 7, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This is a VideoAsPrompt ComfyUI plugin☆21Oct 30, 2025Updated 10 months ago
- InteractAvatar is a novel dual-stream DiT framework that enables talking avatars to perform Grounded Human-Object Interaction (GHOI)☆24Jun 3, 2026Updated 2 months ago
- This is a ComfyUI plugin for https://huggingface.co/spaces/Soul-AILab/SoulX-Singer☆18Feb 12, 2026Updated 6 months ago
- This is a Qwen-Image-i2L ComfyUI plugin☆81Dec 12, 2025Updated 8 months ago
- Comfy EverAnimate WIP☆17Jun 5, 2026Updated 2 months ago
- ☆20Dec 19, 2025Updated 8 months ago
- FlashVSR:Towards Real-Time Diffusion-Based Streaming Video Super-Resolution,you can use it in comfyUI☆392Feb 23, 2026Updated 6 months ago
- LucidFlux: Caption-Free Universal Image Restoration with a Large-Scale Diffusion Transformer,you can use it in ComfyUI☆62May 27, 2026Updated 3 months ago
- enseNova-U1: Unifying Multimodal Understanding and Generation with NEO-Unify Architecture☆87Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ComfyUI custom nodes for LongCat-AudioDiT \ Diffusion-based Zero-Shot Text-to-Speech☆134Apr 4, 2026Updated 4 months ago
- StableAvatar: Infinite-Length Audio-Driven Avatar Video Generation,you can try it in ComfyUI☆44Aug 21, 2025Updated last year
- This is a ComfyUI custom node used to convert Qwen-Image LoRA files trained on the ModelScope platform to a format that ComfyUI can recog…☆30Aug 9, 2025Updated last year
- Implementing FlowEdit, maybe other inversion techniques for the Wan video generation model☆41Feb 28, 2025Updated last year
- interpretability work and exploration for krea☆52Jul 12, 2026Updated last month
- 不限人数的ComfyUI IndexTTS2节点。ComfyUI IndexTTS2 for multiple participants with no limit on the number of speakers.☆34Sep 14, 2025Updated 11 months ago
- ComfyUI custom nodes for Ovi joint video+audio generation☆47Oct 6, 2025Updated 10 months ago
- ComfyUI node for highly expressive speech and realistic zero-shot voice cloning☆509Aug 6, 2026Updated 3 weeks ago
- Wrapper of DyPE: Dynamic Position Extrapolation for Ultra High Resolution Diffusion, run in diffusers mode☆32Feb 26, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆270Jan 4, 2026Updated 7 months ago
- OmniSVG: A Unified Scalable Vector Graphics Generation Model,you can try it in ComfyUI☆31Dec 5, 2025Updated 8 months ago
- ☆58Dec 25, 2025Updated 8 months ago
- A custom ComfyUI node for MiniCPM vision-language models, supporting v4, v4.5, and v4 GGUF formats, enabling high-quality image captionin…☆154Aug 28, 2025Updated last year
- forked from https://github.com/fredconex/ComfyUI-SoundFlow☆25Nov 20, 2025Updated 9 months ago
- Qwen3-TTS在ComfyUI的实现,强大语音生成能力,为语音克隆、语音设计、超高质量类人语音生成以及基于自然语言的语音控制提供全面支持。☆135Jan 27, 2026Updated 7 months ago
- 一款ComfyUI扩展节点,能够为您的图像添加各种精美的艺术文字效果,支持丰富的文字样式和特效。☆30Aug 19, 2026Updated last week
- Quality-preserving per-layer conditioning control for Krea 2 (ComfyUI node). Fork of nova452/ComfyUI-ConditioningKrea2Rebalance — RMS ren…☆144Jun 26, 2026Updated 2 months ago
- ComfyUI自定义节点,集成ByteDance Sa2VA模型,实现智能图像分割和视觉理解。☆29Dec 17, 2025Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A ComfyUI speech recognition plugin based on [Qwen3-ASR](https://github.com/QwenLM/Qwen3-ASR).☆37Feb 6, 2026Updated 6 months ago
- ☆24Sep 20, 2025Updated 11 months ago
- ☆68Dec 12, 2025Updated 8 months ago
- SparkVSR: Interactive Video Super-Resolution via Sparse Keyframe Propagation☆39Mar 31, 2026Updated 4 months ago
- Fun-CineForge: A Unified Dataset Pipeline and Model for Zero-Shot Movie Dubbing in Diverse Cinematic Scenes☆21Apr 2, 2026Updated 4 months ago
- ☆46Oct 26, 2024Updated last year
- 一个高性能的ComfyUI视频&图像放大插件,利用CUDA加速,支持多GPU、混合精度和Tensor Core优化。☆23Oct 26, 2025Updated 10 months ago