👆Pytorch implementation of "Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion"
☆37Jul 28, 2025Updated last year
Alternatives and similar repositories for Ctrl-V
Users that are interested in Ctrl-V are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MagicVFX: Visual Effects Synthesis in Just Minutes☆18Dec 16, 2024Updated last year
- ☆14Oct 16, 2023Updated 2 years ago
- Official repo of "Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without Training“☆28Jan 7, 2026Updated 9 months ago
- Official Pytorch implementation of 'Facing the Elephant in the Room: Visual Prompt Tuning or Full Finetuning'? (ICLR2024)☆13Mar 8, 2024Updated 2 years ago
- [ACM MM24 Poster] Official implementation of paper "MVPbev: Multi-view Perspective Image Generation from BEV with Test-time Controllabili…☆20Sep 6, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR 2025] PoseTraj: Pose-Aware Trajectory Control in Video Diffusion☆23May 26, 2026Updated 4 months ago
- [cvpr2026] UniPR: Unified Object-level Real-to-Sim Perception and Reconstruction from a Single Stereo Pair☆18Mar 27, 2026Updated 6 months ago
- DecMem: Towards Minute-Long Consistent World Generation with Decoupled Memory☆33Jun 1, 2026Updated 4 months ago
- ☆33Jul 5, 2024Updated 2 years ago
- Official Code for DOROTHIE: Spoken Dialogue for Handling Unexpected Situations in Interactive Autonomous Driving Agents (Findings of EMNL…☆22Oct 24, 2023Updated 2 years ago
- DreamHOI: Subject-Driven Generation of 3D Human-Object Interactions with Diffusion Priors☆37Sep 13, 2024Updated 2 years ago
- ☆31Nov 7, 2023Updated 2 years ago
- The official repository of our paper: "Urban Architect: Steerable 3D Urban Scene Generation with Layout Prior"☆113Apr 26, 2024Updated 2 years ago
- This is the official repository for "LatentMan: Generating Consistent Animated Characters using Image Diffusion Models" [CVPRW 2024]☆22Jul 21, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official repository for "Object Wake-up: 3D Object Rigging from a Single Image" (ECCV 2022)☆27Oct 2, 2022Updated 4 years ago
- [ECCV 2024] 3DPE: Real-time 3D-aware Portrait Editing from a Single Image☆22Sep 15, 2025Updated last year
- Official code for "Amodal Completion via Progressive Mixed Context Diffusion" [CVPR 2024 Highlight]☆56Jul 24, 2024Updated 2 years ago
- ☆27Aug 12, 2025Updated last year
- This is the official implementation of SG-I2V: Self-Guided Trajectory Control in Image-to-Video Generation.☆116Nov 26, 2024Updated last year
- Code for Open3DTrack: Towards Open-Vocabulary 3D Multi-Object Tracking☆36Mar 14, 2025Updated last year
- The official code of "Beyond VLM-Based Rewards: Diffusion-Native Latent Reward Modeling"☆68Jun 30, 2026Updated 3 months ago
- [arXiv 2024] I4VGen: Image as Free Stepping Stone for Text-to-Video Generation☆24Oct 6, 2024Updated 2 years ago
- Official pytorch implementation for SingleInsert☆28Apr 19, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- An unofficial reproduction of Sela et al. "Computational caricaturization of surfaces". CVIU 2015.☆13Nov 27, 2020Updated 5 years ago
- ☆11Mar 11, 2024Updated 2 years ago
- Inferring and Leveraging Parts from Object Shape for Improving Semantic Image Synthesis (CVPR 2023)☆18Dec 13, 2024Updated last year
- ☆13Jul 10, 2024Updated 2 years ago
- Official implementation of MCVD: Masked Conditional Video Diffusion for Prediction, Generation, and Interpolation (https://arxiv.org/abs/…☆371Sep 22, 2022Updated 4 years ago
- Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)☆23Mar 18, 2026Updated 6 months ago
- (ICCV 2023) Betrayed by Captions: Joint Caption Grounding and Generation for Open Vocabulary Instance Segmentation☆47Jul 18, 2024Updated 2 years ago
- Modified version of QPBO algorithm by Vladimir Kolmogorov for very large graphs.☆12Dec 14, 2018Updated 7 years ago
- [IJCV 2025] VLPrompt-PSG: Vision-Language Prompting for Panoptic Scene Graph Generation☆27Sep 24, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Implementation of ViLLA-X, Enhancing Latent Action Modeling in Vision-Language-Action Models☆23Aug 27, 2025Updated last year
- Official implementation of "VSTAR: Generative Temporal Nursing for Longer Dynamic Video Synthesis"☆21Jan 26, 2025Updated last year
- [ECCV 2024] Noise Calibration: Plug-and-play Content-Preserving Video Enhancement using Pre-trained Video Diffusion Models☆89Sep 3, 2024Updated 2 years ago
- ☆26Nov 30, 2025Updated 10 months ago
- Combined InstantID🔥 and FouriScale to generate high resolution image!☆11Apr 3, 2024Updated 2 years ago
- [ECCV 2024 Oral] MotionDirector: Motion Customization of Text-to-Video Diffusion Models.☆1,055Aug 21, 2024Updated 2 years ago
- The benchmark for "Video Object Segmentation in Panoptic Wild Scenes".☆13Oct 17, 2023Updated 2 years ago