VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders
☆33Jul 15, 2026Updated this week
Alternatives and similar repositories for VideoRAE
Users that are interested in VideoRAE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft"☆19Oct 11, 2025Updated 9 months ago
- [ICML 2026] "LIVE: Long-horizon Interactive Video World ModEling"☆35Updated this week
- [NeurIPS 2025] Official repo of "Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions"☆20Aug 6, 2025Updated 11 months ago
- [ICCV 2025 Highlight] "Edit360: 2D Image Edits to 3D Assets from Any Angle"☆20Feb 4, 2026Updated 5 months ago
- A simple 3D asset retrieval system based on objaverse. Query any 3D asset using text(CN/EN) or images, inter-modal or cross-modal. Equip…☆16Feb 15, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is temporal repository for 3-D human motion tracking from videos and/or IMUs☆16Apr 9, 2023Updated 3 years ago
- [NeurIPS'25] VidEmo: Affective-Tree Reasoning for Emotion-Centric Video Foundation Models☆15Dec 7, 2025Updated 7 months ago
- Face Generation Work (Preprint)☆21Dec 28, 2024Updated last year
- OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder☆58Jul 14, 2026Updated last week
- [CVPR2026] Exploring Spatial Intelligence from a Generative Perspective☆30Jun 3, 2026Updated last month
- ViGeo: Towards Consistent Video Geometry Estimation☆118Updated this week
- G3Splat: Geometrically Consistent Generalizable Gaussian Splatting☆67Jan 5, 2026Updated 6 months ago
- ☆34Apr 10, 2026Updated 3 months ago
- Project page for SparseSplat: Towards Applicable Feed-Forward 3D Gaussian Splatting with Pixel-Unaligned Prediction (CVPR 2026)☆31Apr 28, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 3 months ago
- Some of my courses fragmentary cheat sheets in LaTex.☆23Sep 26, 2015Updated 10 years ago
- ☆75Jun 22, 2026Updated 3 weeks ago
- ☆26Jul 22, 2025Updated 11 months ago
- [SIGGRAPH 2026 Journal] SegviGen: Repurposing 3D Generative Model for Part Segmentation☆155Mar 19, 2026Updated 4 months ago
- ☆26Jun 2, 2026Updated last month
- Official repository for "Vid2World: Crafting Video Diffusion Models to Interactive World Models" (ICLR 2026), https://arxiv.org/abs/2505.…☆69Jan 27, 2026Updated 5 months ago
- Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation☆15Aug 11, 2025Updated 11 months ago
- Official Implemenation for RAEv2: Improved Baselines with Representation Autoencoders☆308May 21, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Stable-Sim2Real: Exploring Simulation of Real-Captured 3D Data with Two-Stage Depth Diffusion (ICCV 2025 Highlight)☆31Mar 15, 2026Updated 4 months ago
- PyTorch implementation of RiT: Vanilla Diffusion Transformers Suffice in Representation Space☆26May 23, 2026Updated last month
- (ICCV 2025) Holistic Tokenizer for Autoregressive Image Generation☆34Oct 9, 2025Updated 9 months ago
- VideoFlexTok: Flexible-Length Coarse-to-Fine Video Tokenization☆43Apr 16, 2026Updated 3 months ago
- [CVPR 2026] Multi-view Pyramid Transformer: Look Coarser to See Broader☆145Mar 25, 2026Updated 3 months ago
- Metric implementation and raw data of "Diffusing in the Right Space: A Systematic Study of Latent Diffusability"☆34Jun 16, 2026Updated last month
- A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving a…☆659Jun 4, 2026Updated last month
- Context Forcing: Consistent Autoregressive Video Generation with Long Context [ICML26]☆98Jun 29, 2026Updated 3 weeks ago
- The first open-domain closed-loop revisited benchmark for evaluating memory consistency and action control in world models.☆72Jul 2, 2026Updated 2 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [Arxiv'25] DINO-Tok: Adapting DINO for Visual Tokenizers☆40Apr 11, 2026Updated 3 months ago
- Official Implementation of Paper FOLDER (ICCV2025) and Turbo (ECCV2024)☆15Jun 27, 2025Updated last year
- ☆94Apr 29, 2026Updated 2 months ago
- ☆12Mar 25, 2024Updated 2 years ago
- The official codes of Learning to Decouple the Lights for 3D Face Texture Modeling (NeurIPS'24)☆14Mar 17, 2025Updated last year
- A toolbox for processing Total Capture dataset☆32Apr 10, 2020Updated 6 years ago
- Various functions I found useful for prototyping computer graphics code in Python☆11Dec 27, 2019Updated 6 years ago