Kaleido: Open-sourced multi-subject reference video generation model, enabling controllable, high-fidelity video synthesis from multiple image references.
☆148Mar 2, 2026Updated 5 months ago
Alternatives and similar repositories for Kaleido
Users that are interested in Kaleido are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement (ICLR2026)☆299Mar 24, 2026Updated 5 months ago
- A real-time streaming conversational video system that transforms text interactions into continuous, high-fidelity video responses using …☆340Dec 15, 2025Updated 8 months ago
- Concat-ID: Towards Universal Identity-Preserving Video Synthesis☆65May 7, 2025Updated last year
- ☆15Oct 27, 2025Updated 10 months ago
- ☆15Jun 21, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Pose Extraction & Rendering for SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representat…☆229Jun 11, 2026Updated 2 months ago
- [CVPR2026] Official implementation of our paper “Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot…☆20Apr 8, 2026Updated 4 months ago
- [CVPR 2026] Scaling Zero-Shot Reference-to-Video Generation☆77Apr 28, 2026Updated 4 months ago
- [ECCV 2026] ShotStream: Streaming Multi-Shot Video Generation for Interactive Storytelling☆181Jun 23, 2026Updated 2 months ago
- Video Content Customization Using First Frame☆194Mar 17, 2026Updated 5 months ago
- ☆52Jan 6, 2026Updated 7 months ago
- [ICLR-26, ECCV-26, NeurIPS-25] Lumos-Custom Project: research for customized video generation in the Lumos Project.☆221Jun 29, 2026Updated 2 months ago
- HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning☆1,284Jan 25, 2026Updated 7 months ago
- CVPR 2026 | Official Implementation of "MultiShotMaster: A Controllable Multi-Shot Video Generation Framework"☆177Feb 22, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Phantom: Subject-Consistent Video Generation via Cross-Modal Alignment☆1,516Sep 11, 2025Updated 11 months ago
- [SIGGRAPH Asia'25] Enabling Reference-based Camera Control via Context without Explicit 3D Estimation☆160Jan 18, 2026Updated 7 months ago
- SCAIL: Towards Studio-Grade Character Animation via In-Context Learning of 3D-Consistent Pose Representations (CVPR 2026 Findings)☆1,045May 6, 2026Updated 3 months ago
- Official code for StoryMem: Multi-shot Long Video Storytelling with Memory☆762Jul 22, 2026Updated last month
- [ICML 2026] DreamID-Omni: Unified Framework for Controllable Human-Centric Audio-Video Generation☆277May 22, 2026Updated 3 months ago
- [ICLR 2026] Official Repo For "BindWeave: Subject-Consistent Video Generation via Cross-Modal Integration"☆339Jan 28, 2026Updated 7 months ago
- Mixture-of-Groups Attention for End-to-End Long Video Generation☆100Oct 22, 2025Updated 10 months ago
- [AAAI 2026] FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation☆65Aug 20, 2025Updated last year
- [CVPR 2026 Highlight] Official Implementations for Paper - HoloCine: Holistic Generation of Cinematic Multi-Shot Long Video Narratives☆700Nov 26, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A unified and fully open-source framework for instruction-guided and reference-guided video editing using natural language.☆319May 13, 2026Updated 3 months ago
- GLM-Image: Auto-regressive for Dense-knowledge and High-fidelity Image Generation.☆1,022Mar 20, 2026Updated 5 months ago
- ☆88Nov 16, 2025Updated 9 months ago
- [NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation☆780Apr 16, 2026Updated 4 months ago
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆167Jun 26, 2026Updated 2 months ago
- [ICLR 2026] Lumos Project: Frontier video unified model research by Alibaba DAMO Academy.☆262Apr 6, 2026Updated 4 months ago
- [NeurIPS 2025] Improving Video Generation with Human Feedback☆501Sep 24, 2025Updated 11 months ago
- [CVPR 2026 Highlight] Official implementation of BiCo: Composing Concepts from Images and Videos via Concept-prompt Binding☆86May 31, 2026Updated 3 months ago
- OmniInsert: Mask-Free Video Insertion of Any Reference via Diffusion Transformer Models☆162Mar 4, 2026Updated 5 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Phantom-Data: Towards a General Subject-Consistent Video Generation Dataset☆120Feb 25, 2026Updated 6 months ago
- ☆56May 6, 2026Updated 3 months ago
- [ECCV 2026] Official Code of "Distribution Matching Distillation Meets Reinforcement Learning"☆293Feb 1, 2026Updated 6 months ago
- ☆61Dec 1, 2025Updated 8 months ago
- [ICLR 2026] Official repo for paper "Video-As-Prompt: Unified Semantic Control for Video Generation"☆450Feb 8, 2026Updated 6 months ago
- Blending Custom Photos with Video Diffusion Transformers☆50Jan 21, 2025Updated last year
- Bernini is a unified framework for video generation and editing that combines an MLLM-based semantic planner with a DiT-based renderer.☆1,292Aug 13, 2026Updated 2 weeks ago