Evolution of Video Generative Foundations
☆47Apr 7, 2026Updated 5 months ago
Alternatives and similar repositories for Awesome-Video-Foundations
Users that are interested in Awesome-Video-Foundations are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [CVPR25] IAR☆18Jun 13, 2025Updated last year
- Official implementation for "DiTraj: Training-free Trajectory Control For Video Diffusion Transformer"☆21Dec 11, 2025Updated 9 months ago
- Audio-video joint generation☆58Nov 27, 2025Updated 9 months ago
- This is a collective repository for all 3D and 4D Object Generation papers☆20May 22, 2026Updated 3 months ago
- Video Diffusion Transformers are In-Context Learners☆37Jan 6, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [TMM] MINT-IQA: Quality Assessment for AI Generated Images with Instruction Tuning☆21Nov 21, 2025Updated 9 months ago
- [ICCV2025] The official code of "DreamRelation: Relation-Centric Video Customization"☆26Feb 4, 2026Updated 7 months ago
- [ArXiv 2025] A survey about controllable video generation: This repo is the official awesome of "Controllable video generation: A survey…☆772Jul 31, 2026Updated last month
- ☆39Jul 27, 2024Updated 2 years ago
- [ICML 2026] The official implementation of paper "Unified Multimodal Autoregressive Modeling with Shared Context—Visual Tokenizer is Key …☆58Jul 13, 2026Updated 2 months ago
- PexelsCustom-1M: A Comprehensive Ecosystem for Open-Domain Customized Video Generation☆19Jun 30, 2026Updated 2 months ago
- [CVPR2026] Official implementation of our paper “Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot…☆20Apr 8, 2026Updated 5 months ago
- [ICLR26] Understanding VS. Generation: Navigating Optimization Dilemma in Multimodal Models☆28May 6, 2026Updated 4 months ago
- UniAVGen: Unified Audio and Video Generation with Asymmetric Cross-Modal Interactions☆59Dec 16, 2025Updated 9 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Code as World: Agentic Discovery of Executable World Representations for Physical Reasoning☆452Aug 31, 2026Updated 2 weeks ago
- [CVPR 2026] Video-as-Answer: Predict and Generate Next Video Event with Joint-GRPO☆119Feb 28, 2026Updated 6 months ago
- ☆36Feb 18, 2026Updated 7 months ago
- Code for "Skill-based Chain-of-Thoughts for Domain-Adaptive Video Reasoning [EMNLP 2025 Findings]"☆18Aug 27, 2025Updated last year
- ☆41Jan 21, 2026Updated 7 months ago
- Developer project for getting basic API integrations working in under 5 minutes☆11May 22, 2026Updated 3 months ago
- Official PyTorch implementation of paper “InsViE-1M: Effective Instruction-based Video Editing with Elaborate Dataset Construction”☆35Apr 3, 2026Updated 5 months ago
- [ICCV 25] VMBench: A Benchmark for Perception-Aligned Video Motion Generation☆76Oct 10, 2025Updated 11 months ago
- [NIPS2025] RoPECraft: Training-Free Motion Transfer with Trajectory-Guided RoPE Optimization on Diffusion Transformers☆52Sep 24, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official Implementation for "SiLVR : A Simple Language-based Video Reasoning Framework"☆20Jan 18, 2026Updated 8 months ago
- Project page of "ICLR-2026 Echo: Towards Advanced Audio Comprehension via Audio-Interleaved Reasoning"☆19Mar 26, 2026Updated 5 months ago
- Official implementation of "Conditional Score Guidance for Text-Driven Image-to-Image Translation" (NeurIPS 2023)☆11Jul 19, 2023Updated 3 years ago
- [AAAI 2026] PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching☆28Feb 4, 2026Updated 7 months ago
- ☆50Jan 15, 2026Updated 8 months ago
- OpenVE-3M: A Large-Scale High-Quality Dataset for Instruction-Guided Video Editing☆53Apr 15, 2026Updated 5 months ago
- Code for paper "RapVerse: Coherent Vocals and Whole-Body Motions Generations from Text"☆18May 30, 2024Updated 2 years ago
- ☆17Mar 19, 2026Updated 6 months ago
- [ICML 2026] Transform Trained Transformer for Accelerating Native 4K Video Generation☆41Dec 16, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement☆21Aug 4, 2026Updated last month
- [CVPR 2025] GPS as a Control Signal for Image Generation☆25Mar 18, 2025Updated last year
- The official code implementation of Generalized Category Discovery in Semantic Segmentation☆17Dec 20, 2023Updated 2 years ago
- [CVPR 2026 Highlight🎉] Official implementation of WorldForge☆177Apr 10, 2026Updated 5 months ago
- ☆17Mar 24, 2026Updated 5 months ago
- [T-PAMI 2025] EMOv2: Pushing 5M Vision Model Frontier☆54Dec 30, 2024Updated last year
- NEWTON: Agentic Planning for Physically Grounded Video Generation☆143Jul 31, 2026Updated last month