☆69Jul 31, 2026Updated this week
Alternatives and similar repositories for InstructAV2AV
Users that are interested in InstructAV2AV are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of "SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers"☆68Jul 4, 2026Updated 3 weeks ago
- Code Release for "OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data"☆78Jun 12, 2026Updated last month
- ☆108May 27, 2026Updated 2 months ago
- [arXiv 2026] Project page for paper "SwiftI2V: Efficient High-Resolution Image-to-Video Generation via Conditional Segment-wise Generatio…☆86May 8, 2026Updated 2 months ago
- SCOPE: Simulating Cross-game Operations in Playable Environments for FPS World Models☆74May 28, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆23Oct 15, 2025Updated 9 months ago
- open source style transfer model on par with nano banana pro, supporting Qwen-Image-Edit 2509, 2511, SenseNovaU1☆96Updated this week
- PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion☆1,007Jul 22, 2026Updated last week
- ☆90May 13, 2026Updated 2 months ago
- [ICML 2026] Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization☆64Jun 11, 2026Updated last month
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆164Jun 26, 2026Updated last month
- The Pulse of Motion: Measuring Physical Frame Rate from Visual Dynamics☆73Mar 26, 2026Updated 4 months ago
- MTVCraft: An Open Veo3-style Audio-Video Generation Demo☆98Oct 8, 2025Updated 9 months ago
- PIXLRelight: Controllable Relighting via Intrinsic Conditioning☆69May 20, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [Official Code] PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory☆43Jun 17, 2026Updated last month
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆214Jun 30, 2026Updated last month
- Official Repo For PerceptionDLM Codebase☆77Jun 22, 2026Updated last month
- Implementation for for "L-CoDer: Language-based Colorization with Color-object Decoupling Transformer"☆13Jan 20, 2024Updated 2 years ago
- LOGOS (Language Of Generative Objects in Science) is the first multi-domain generative foundation model for the natural sciences built on…☆139Jun 25, 2026Updated last month
- 【SIGGRAPH Asia 2026】Official repo for the paper "PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesi…☆153Jun 22, 2026Updated last month
- ☆110Sep 3, 2025Updated 11 months ago
- ☆104Mar 13, 2026Updated 4 months ago
- Official implementation and project page of the CVPR'24 paper "VMINer: Versatile Multi-view Inverse Rendering with Near- and Far-field Li…☆14Aug 6, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation for "L-CoIns: Language-based Colorization with Instance Awareness"☆11Dec 7, 2023Updated 2 years ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)☆119Aug 15, 2025Updated 11 months ago
- PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects☆306Jun 11, 2026Updated last month
- Implementation of <Streaming Autoregressive Video Generation via Diagonal Distillation> in ICLR 2026☆129Mar 18, 2026Updated 4 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆350Jun 24, 2026Updated last month
- ☆83Oct 13, 2025Updated 9 months ago
- Official repository for "PanoWan: Lifting Diffusion Video Generation Models to 360° with Latitude/Longitude-aware Mechanisms"☆52Dec 18, 2025Updated 7 months ago
- ☆17Apr 23, 2024Updated 2 years ago
- Code for 'JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion'☆268May 11, 2026Updated 2 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆125Apr 30, 2026Updated 3 months ago
- ☆150Jul 24, 2026Updated last week
- UniMesh: Unifying 3D Mesh Understanding and Generation☆57Jul 14, 2026Updated 3 weeks ago
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Models☆96Jun 25, 2026Updated last month
- [ECCV 2026 Oral] Official implementation of "OmniForcing: Unleashing Real-time Joint Audio-Visual Generation"[arXiv:2603.11647]. OmniForc…☆173Jul 23, 2026Updated last week
- Official Implementation of SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning☆1,071Jul 16, 2026Updated 2 weeks ago
- A curated and auto-updated collection of video diffusion / video generation papers from arXiv, covering text-to-video, image-to-video, co…☆31Jul 20, 2026Updated 2 weeks ago