Kaleido: Open-sourced multi-subject reference video generation model, enabling controllable, high-fidelity video synthesis from multiple image references.
☆149Mar 2, 2026Updated 6 months ago
Alternatives and similar repositories for Kaleido
Users that are interested in Kaleido are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement (ICLR2026)☆299Mar 24, 2026Updated 5 months ago
- A real-time streaming conversational video system that transforms text interactions into continuous, high-fidelity video responses using …☆343Dec 15, 2025Updated 9 months ago
- Concat-ID: Towards Universal Identity-Preserving Video Synthesis☆65May 7, 2025Updated last year
- ☆15Oct 27, 2025Updated 10 months ago
- ☆15Jun 21, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [NeurIPS 2025 D&B🔥] OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation☆230May 19, 2026Updated 4 months ago
- Pose Extraction & Rendering for SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representat…☆230Jun 11, 2026Updated 3 months ago
- [CVPR2026] Official implementation of our paper “Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot…☆20Apr 8, 2026Updated 5 months ago
- [CVPR 2026] Scaling Zero-Shot Reference-to-Video Generation☆77Apr 28, 2026Updated 4 months ago
- [ECCV 2026] ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling☆181Updated this week
- Video Content Customization Using First Frame☆194Mar 17, 2026Updated 6 months ago
- ☆52Jan 6, 2026Updated 8 months ago
- [ICLR-26, ECCV-26, NeurIPS-25] Lumos-Custom Project: research for customized video generation in the Lumos Project.☆222Jun 29, 2026Updated 2 months ago
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning☆1,287Jan 25, 2026Updated 7 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- CVPR 2026 | Official Implementation of "MultiShotMaster: A Controllable Multi-Shot Video Generation Framework"☆178Feb 22, 2026Updated 6 months ago
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment☆1,516Sep 11, 2025Updated last year
- [SIGGRAPH Asia'25] Enabling Reference-based Camera Control via Context without Explicit 3D Estimation☆160Jan 18, 2026Updated 8 months ago
- SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations (CVPR 2026 Findings)☆1,053May 6, 2026Updated 4 months ago
- Official code for StoryMem: Multi-shot Long Video Storytelling with Memory☆769Jul 22, 2026Updated last month
- [ICML 2026] DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation☆281May 22, 2026Updated 3 months ago
- [ICLR 2026] Official Repo For "BindWeave: Subject-Consistent Video Generation via Cross-Modal Integration"☆340Jan 28, 2026Updated 7 months ago
- Mixture-of-Groups Attention for End-to-End Long Video Generation☆100Oct 22, 2025Updated 10 months ago
- [AAAI 2026] FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation☆65Aug 20, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2026 Highlight] Official Implementations for Paper - HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives☆703Nov 26, 2025Updated 9 months ago
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.☆325May 13, 2026Updated 4 months ago
- GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image Generation.☆1,037Mar 20, 2026Updated 6 months ago
- ☆87Nov 16, 2025Updated 10 months ago
- [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation☆786Apr 16, 2026Updated 5 months ago
- [ICLR 2026] Lumos Project: Frontier video unified model research by Alibaba DAMO Academy.☆260Apr 6, 2026Updated 5 months ago
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆168Jun 26, 2026Updated 2 months ago
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆511Sep 24, 2025Updated 11 months ago
- [CVPR 2026 Highlight] Official implementation of BiCo: Composing Concepts from Images and Videos via Concept-prompt Binding☆87May 31, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models☆162Mar 4, 2026Updated 6 months ago
- Phantom-Data: Towards a General Subject-Consistent Video Generation Dataset☆120Feb 25, 2026Updated 6 months ago
- ☆56May 6, 2026Updated 4 months ago
- [ECCV 2026] Official Code of "Distribution Matching Distillation Meets Reinforcement Learning"☆299Feb 1, 2026Updated 7 months ago
- ☆60Dec 1, 2025Updated 9 months ago
- [ICLR 2026] Official repo for paper "Video-As-Prompt: Unified Semantic Control for Video Generation"☆454Feb 8, 2026Updated 7 months ago
- Blending Custom Photos with Video Diffusion Transformers☆50Jan 21, 2025Updated last year