Learning from Next-Frame Prediction: Autoregressive Video Modeling Encodes Effective Representations
☆22Dec 24, 2025Updated 8 months ago
Alternatives and similar repositories for NExT-Vid
Users that are interested in NExT-Vid are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆12May 5, 2025Updated last year
- Official codes for the paper "GARDO: Reinforcing Diffusion Models without Reward Hacking"☆63May 3, 2026Updated 4 months ago
- [CVPR-26] Official repository of "CaTok: Taming Mean Flows for One-Dimensional Causal Image Tokenization"☆19Mar 9, 2026Updated 6 months ago
- Code release for "Memorization in 3D Shape Generation: An Empirical Study"☆21Dec 30, 2025Updated 8 months ago
- Codebase for the paper-Elucidating the design space of language models for image generation☆44Nov 17, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Minute-long video generation at 24FPS.☆70Mar 28, 2026Updated 5 months ago
- ☆38Dec 25, 2025Updated 8 months ago
- RationalRewards: a reasoning reward model for diffusion RL and test-time prompt tuning☆61Jun 4, 2026Updated 3 months ago
- [ICIP 2025] Scribble-Guided Diffusion for Training-free Text-to-Image Generation☆27Oct 2, 2024Updated last year
- ☆34Dec 29, 2025Updated 8 months ago
- [NeurIPS 2024] ReNO: Enhancing One-step Text-to-Image Models through Reward-based Noise Optimization☆167Sep 15, 2025Updated last year
- ☆39Feb 24, 2026Updated 6 months ago
- Official inference code and LongText-Bench benchmark for our paper X-Omni (https://arxiv.org/pdf/2507.22058).☆428Aug 26, 2025Updated last year
- [NeurIPS'25 Spotlight] MJ-VIDEO: Fine-Grained Benchmarking and Rewarding Video Preferences in Video Generation☆20Feb 23, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official Code for Talk2Move: Reinforcement Learning for Text-Instructed Object-Level Geometric Transformation in Scenes.☆25Apr 3, 2026Updated 5 months ago
- This repository provides the official implementation of VTBench, a benchmark designed to evaluate the performance of visual tokenizers (V…☆36Jul 30, 2025Updated last year
- PICABench: How Far Are We from Physically Realistic Image Editing?☆41Nov 5, 2025Updated 10 months ago
- SegLocNet: Multimodal Localization Network for Autonomous Driving via Bird’s-Eye-View Segmentation☆33May 9, 2025Updated last year
- Developer project for getting basic API integrations working in under 5 minutes☆11May 22, 2026Updated 3 months ago
- [CVPR 2026] Official repo for "EVATok: Adaptive Length Video Tokenization for Efficient Visual Autoregressive Generation"☆65Mar 13, 2026Updated 6 months ago
- [CVPR 2023] Out-of-Distributed Semantic Pruning for Robust Semi-Supervised Learning☆22Jun 11, 2023Updated 3 years ago
- [NeurIPS 2024] RectifID: Personalizing Rectified Flow with Anchored Classifier Guidance☆130Oct 13, 2024Updated last year
- SpotEdit [NeurIPS 2025 W]☆18Sep 24, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official Pytorch implementation for LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior (ICLR 2025 Oral).☆108Feb 11, 2025Updated last year
- [CVPR 2026🔥] Enhancing Spatial Understanding in Image Generation via Reward Modeling☆86Mar 2, 2026Updated 6 months ago
- ☆49Feb 20, 2026Updated 6 months ago
- Official Repository of Personalized Visual Instruct Tuning☆34Mar 6, 2025Updated last year
- RePO: Replay-Enhanced Policy Optimization☆24Jun 12, 2025Updated last year
- Official implementation of Add-SD: Rational Generation without Manual Reference.☆29Aug 19, 2024Updated 2 years ago
- [CVPR 2026 (Highlight)] Unofficial Implementation of "Image Diffusion Preview with Consistency Solver"☆31Jan 24, 2026Updated 7 months ago
- ☆32Jul 25, 2026Updated last month
- [AAAI 2026] SlideTailor: Personalized Presentation Slide Generation for Scientific Papers☆59Apr 18, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repo for paper "IC-Effect: Precise and Efficient Video Effects Editing via In-Context Learning"☆43Jan 29, 2026Updated 7 months ago
- ☆13Jul 5, 2024Updated 2 years ago
- [ICML 2026] Official repo for "DiffThinker: Towards Generative Multimodal Reasoning with Diffusion Models"☆186Jan 4, 2026Updated 8 months ago
- [ECCV 2024] 3DPE: Real-time 3D-aware Portrait Editing from a Single Image☆22Sep 15, 2025Updated last year
- [ICLR2026] The code for "Interp3D: Correspondence-Aware Interpolation for Generative Textured 3D Morphing."☆33Jan 21, 2026Updated 7 months ago
- (ICCV 2023) MasQCLIP for Open-Vocabulary Universal Image Segmentation☆37Oct 18, 2023Updated 2 years ago
- [NeurIPS 2024] Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective☆78Oct 31, 2024Updated last year