[ICML 2026] DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation
☆281May 22, 2026Updated 4 months ago
Alternatives and similar repositories for DreamID-Omni
Users that are interested in DreamID-Omni are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV 2026 Oral] DreamID-V: Bridging the Image-to-Video Gap for High-Fidelity Face Swapping via Diffusion Transformer☆684May 22, 2026Updated 4 months ago
- ☆17Mar 19, 2026Updated 6 months ago
- OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models☆162Mar 4, 2026Updated 6 months ago
- OmniTransfer: All-in-one Framework for Spatio-temporal Video Transfer☆236Apr 15, 2026Updated 5 months ago
- [CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length vide…☆481Feb 21, 2026Updated 7 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.☆327May 13, 2026Updated 4 months ago
- Official Implementation of CoInteract: Spatially-Structured Co-Generation for Interactive Human-Object Video Synthesis☆170May 7, 2026Updated 4 months ago
- A foundation model that generates synchronized video and audio in a single model☆1,117Updated this week
- DreamStyle: A Unified Framework for Video Stylization☆124Jan 7, 2026Updated 8 months ago
- [ICLR-26, ECCV-26, NeurIPS-25] Lumos-Custom Project: research for customized video generation in the Lumos Project.☆222Jun 29, 2026Updated 2 months ago
- [ECCV 2026] ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling☆181Sep 16, 2026Updated last week
- ☆110Jan 6, 2026Updated 8 months ago
- [ICML2026] From Statics to Dynamics: Physics-Aware Image Editing with Latent Transition Priors☆93Apr 30, 2026Updated 4 months ago
- [ECCV 2026 Oral] Official implementation of "OmniForcing: Unleashing Real-time Joint Audio-Visual Generation"[arXiv:2603.11647]. OmniForc…☆195Jul 23, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [CVPR 2026 Highlight] VideoCoF: Unified Video Editing with Temporal Reasoner☆210Jun 17, 2026Updated 3 months ago
- ☆104Mar 13, 2026Updated 6 months ago
- [CVPR 2026 Poster] One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer☆498Apr 19, 2026Updated 5 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆359Jun 24, 2026Updated 2 months ago
- [ICML 2026] ByteDance's All-in-One Video Generation Model for Human-Object Interaction Video Generation☆472Jul 29, 2026Updated last month
- ☆2,117Apr 11, 2026Updated 5 months ago
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning☆1,286Jan 25, 2026Updated 7 months ago
- Phantom-Data: Towards a General Subject-Consistent Video Generation Dataset☆120Feb 25, 2026Updated 6 months ago
- UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions☆59Dec 16, 2025Updated 9 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆225Jun 30, 2026Updated 2 months ago
- [CVPR2026 🎉] Stand-In is a lightweight, plug-and-play framework for identity-preserving video generation.☆791Aug 10, 2026Updated last month
- We propose a novel modular framework that learns to dynamically mix low-rank adapters (LoRAs) to improve visual analogy learning, enablin…☆76Aug 2, 2026Updated last month
- [CVPR 2026] Scaling Zero-Shot Reference-to-Video Generation☆77Apr 28, 2026Updated 4 months ago
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment☆1,517Sep 11, 2025Updated last year
- Kaleido: Open-sourced multi-subject reference video generation model, enabling controllable, high-fidelity video synthesis from multiple …☆149Mar 2, 2026Updated 6 months ago
- FantasyPortrait: Enhancing Multi-Character Portrait Animation with Expression-Augmented Diffusion Transformers☆509Aug 20, 2025Updated last year
- ☆86Oct 10, 2025Updated 11 months ago
- Accepted by ICML2026☆94Jun 29, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Helios: Real Real-Time Long Video Generation Model☆2,158Aug 24, 2026Updated 3 weeks ago
- [CVPR'26] VecGlypher: Unified Vector Glyph Generation with Language Models☆148Feb 26, 2026Updated 6 months ago
- ☆93May 13, 2026Updated 4 months ago
- ☆1,760Nov 15, 2025Updated 10 months ago
- Official Pytorch Implementation for "Time-to-Move: Training-Free Motion Controlled Video Generation via Dual-Clock Denoising"☆373May 12, 2026Updated 4 months ago
- MoCha: End-to-End Video Character Replacement without Structural Guidance☆784Apr 30, 2026Updated 4 months ago
- [AAAI 2026] UltraGen☆79Feb 1, 2026Updated 7 months ago