VideoDeltaNet-H3: Live T2VA / I2VA / FL2VA / Ref2VA-like generation based on Minimax H3.
☆544Sep 24, 2026Updated this week
Alternatives and similar repositories for vdn-minimax-h3
Users that are interested in vdn-minimax-h3 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Nov 28, 2023Updated 2 years ago
- [CVPR 2026] Scaling Zero-Shot Reference-to-Video Generation☆77Apr 28, 2026Updated 4 months ago
- Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence☆962Sep 17, 2026Updated last week
- Multi-segment Bernini Director for official ComfyUI Bernini-R☆101Jul 29, 2026Updated last month
- [ICML 2026] Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization☆64Jun 11, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [Neurips 2026] PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory☆44Jun 17, 2026Updated 3 months ago
- Code of the paper "FreePCA:Integrating Consistency Information across Long-short Frames in Training-free Long Video Generation via Princi…☆27Apr 3, 2026Updated 5 months ago
- [SIGGRAPH Asia 2026] UniMate: One Unified Model to Animate Diverse Skeletons☆256Sep 11, 2026Updated 2 weeks ago
- PyTorch implementation of paper "Multi-Modal Proxy Learning Towards Personalized Visual Multiple Clustering" (CVPR 2024)☆15Jul 17, 2024Updated 2 years ago
- Kaleido: Open-sourced multi-subject reference video generation model, enabling controllable, high-fidelity video synthesis from multiple …☆149Mar 2, 2026Updated 6 months ago
- ☆15Mar 30, 2025Updated last year
- Official PyTorch Implementation of Paper "Hand2World: Autoregressive Egocentric Interaction Generation via Free-Space Hand Gestures"☆29Jun 30, 2026Updated 2 months ago
- ☆40Jun 23, 2026Updated 3 months ago
- A Unified Visual Generator with Interleaved OmniModal Context☆235Mar 5, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official PyTorch Implementation of Ctrl-Crash 💥☆54Jun 3, 2025Updated last year
- ☆56May 6, 2026Updated 4 months ago
- [ICML 2026] Code2Worlds: Empowering Coding LLMs for 4D World Generation☆130Jun 3, 2026Updated 3 months ago
- ☆26Feb 6, 2026Updated 7 months ago
- [Arxiv 2025] Official PyTorch Implementation of "SVG-T2I: Scaling up Text-to-Image Latent Diffusion Model Without Variational Autoencoder…☆154Dec 18, 2025Updated 9 months ago
- Official implementation of MAGREF: Masked Guidance for Any-Reference Video Generation with Subject Disentanglement (ICLR2026)☆299Mar 24, 2026Updated 6 months ago
- ☆86Jul 3, 2026Updated 2 months ago
- An inference-time, plug-and-play method for temporal control in multi-event generation☆200Sep 11, 2026Updated 2 weeks ago
- [ICML2026] Official Implementation of "TAG: Tangential Amplifying Guidance for Hallucination-Resistant Sampling"☆45Jul 6, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code Release for "OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data"☆80Aug 4, 2026Updated last month
- [NeurIPS 2026] Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆225Jun 30, 2026Updated 2 months ago
- [ICLR 2025] Where Am I and What Will I See : An Auto-Regressive Model for Spatial Localization and View Prediction☆45Aug 9, 2025Updated last year
- Continuous-Time Distribution Matching for Few-Step Diffusion Distillation👏☆155May 11, 2026Updated 4 months ago
- Infinite Worlds with Versatile Interactions☆1,822Sep 10, 2026Updated 2 weeks ago
- Interactively Segment 3d Meshes, with your clicking and SAM3.☆19Feb 13, 2026Updated 7 months ago
- Unified World Model Inference & Evaluation Infrastructure☆318Sep 10, 2026Updated 2 weeks ago
- ☆83Oct 13, 2025Updated 11 months ago
- ☆27Feb 10, 2026Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated last year
- Official repo for paper "Echo-Infinity: Learnable Evolving Memory for Real-Time Infinite Video Generation"☆113Jun 4, 2026Updated 3 months ago
- [ECCV2026] Official code for MoCam: Unified Novel View Synthesis via Structured Denoising Dynamics☆68May 13, 2026Updated 4 months ago
- [NeurIPS 2025 D&B🔥] OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation☆231May 19, 2026Updated 4 months ago
- UniMesh: Unifying 3D Mesh Understanding and Generation☆57Jul 14, 2026Updated 2 months ago
- [CVPR 2026🔥] Enhancing Spatial Understanding in Image Generation via Reward Modeling☆86Mar 2, 2026Updated 6 months ago
- Real-Time Physical Action-Conditioned Video Generation☆227Mar 6, 2026Updated 6 months ago