☆156Jul 31, 2026Updated last month
Alternatives and similar repositories for InstructAV2AV
Users that are interested in InstructAV2AV are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of "SEGA: Spectral-Energy Guided Attention for Resolution Extrapolation in Diffusion Transformers"☆72Jul 4, 2026Updated 2 months ago
- Code Release for "OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data"☆81Aug 4, 2026Updated last month
- ☆108May 27, 2026Updated 3 months ago
- [arXiv 2026] Project page for paper "SwiftI2V: Efficient High-Resolution Image-to-Video Generation via Conditional Segment-wise Generatio…☆85May 8, 2026Updated 4 months ago
- SCOPE: Simulating Cross-game Operations in Playable Environments for FPS World Models☆79May 28, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆23Oct 15, 2025Updated 10 months ago
- open source style transfer model on par with nano banana pro, supporting Qwen-Image-Edit 2509, 2511, SenseNovaU1☆107Aug 17, 2026Updated 3 weeks ago
- PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion☆1,061Jul 22, 2026Updated last month
- ☆92May 13, 2026Updated 4 months ago
- [ICML 2026] Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization☆64Jun 11, 2026Updated 3 months ago
- DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation☆168Jun 26, 2026Updated 2 months ago
- The Pulse of Motion: Measuring Physical Frame Rate from Visual Dynamics☆74Mar 26, 2026Updated 5 months ago
- MTVCraft: An Open Veo3-style Audio-Video Generation Demo☆99Oct 8, 2025Updated 11 months ago
- PIXLRelight: Controllable Relighting via Intrinsic Conditioning☆72May 20, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [Official Code] PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory☆44Jun 17, 2026Updated 2 months ago
- Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆226Jun 30, 2026Updated 2 months ago
- Official Repo For PerceptionDLM Codebase☆78Jun 22, 2026Updated 2 months ago
- Implementation for for "L-CoDer: Language-based Colorization with Color-object Decoupling Transformer"☆13Jan 20, 2024Updated 2 years ago
- 【SIGGRAPH Asia 2026】Official repo for the paper "PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesi…☆168Aug 4, 2026Updated last month
- LOGOS (Language Of Generative Objects in Science) is the first multi-domain generative foundation model for the natural sciences built on…☆146Jun 25, 2026Updated 2 months ago
- ☆110Sep 3, 2025Updated last year
- ☆104Mar 13, 2026Updated 6 months ago
- Official implementation and project page of the CVPR'24 paper "VMINer: Versatile Multi-view Inverse Rendering with Near- and Far-field Li…☆14Aug 6, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Implementation for "L-CoIns: Language-based Colorization with Instance Awareness"☆11Dec 7, 2023Updated 2 years ago
- Official Implementation of DRA-Ctrl (Dimension-Reduction Attack! Video Generative Models are Experts on Controllable Image Synthesis)☆119Aug 15, 2025Updated last year
- PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects☆316Jun 11, 2026Updated 3 months ago
- Implementation of <Streaming Autoregressive Video Generation via Diagonal Distillation> in ICLR 2026☆132Mar 18, 2026Updated 5 months ago
- [ECCV 2026] Generate high resolution videos with a custom voice and appearance, based on LTX-2/LTX-2.3 + Identity In-Context LoRA☆353Jun 24, 2026Updated 2 months ago
- ☆83Oct 13, 2025Updated 10 months ago
- Official repository for "PanoWan: Lifting Diffusion Video Generation Models to 360° with Latitude/Longitude-aware Mechanisms"☆54Dec 18, 2025Updated 8 months ago
- ☆17Apr 23, 2024Updated 2 years ago
- Code for 'JUST-DUB-IT: Video Dubbing via Joint Audio-Visual Diffusion'☆267May 11, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2026] Adaptive Spectral Feature Forecasting for Diffusion Sampling Acceleration☆134Apr 30, 2026Updated 4 months ago
- ☆157Jul 24, 2026Updated last month
- UniMesh: Unifying 3D Mesh Understanding and Generation☆57Jul 14, 2026Updated last month
- [ECCV 2026] CustomX: Unified Character, Action, and Scene Customization in Video World Models☆97Sep 3, 2026Updated last week
- [ECCV 2026 Oral] Official implementation of "OmniForcing: Unleashing Real-time Joint Audio-Visual Generation"[arXiv:2603.11647]. OmniForc…☆194Jul 23, 2026Updated last month
- Official Implementation of SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning☆1,184Aug 24, 2026Updated 2 weeks ago
- A curated and auto-updated collection of video diffusion / video generation papers from arXiv, covering text-to-video, image-to-video, co…☆41Updated this week