VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders
☆46Jul 15, 2026Updated 3 weeks ago
Alternatives and similar repositories for VideoRAE
Users that are interested in VideoRAE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Memory Forcing: Spatio-Temporal Memory for Consistent Scene Generation on Minecraft"☆19Oct 11, 2025Updated 9 months ago
- [ICML 2026] "LIVE: Long-horizon Interactive Video World ModEling"☆38Jul 15, 2026Updated 3 weeks ago
- official code of Efficient Depth-Guided Urban View Synthesis☆14Dec 24, 2024Updated last year
- 3D semantic segmentation generation built on TRELLIS.2. Better generalization achieved by leveraging 2D segmentation priors when trainin…☆24Jul 23, 2026Updated 2 weeks ago
- The first "ImageNet" 3D dataset.☆90Jul 26, 2026Updated 2 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆38May 19, 2026Updated 2 months ago
- [ICML 2026] Official code for paper: Test-Time Training with KV Binding Is Secretly Linear Attention☆39Apr 30, 2026Updated 3 months ago
- Official implementation of “Signal Structure-Aware Gaussian Splatting for Large-Scale Scene Reconstruction” (ICLR 2026).☆20Jul 4, 2026Updated last month
- ☆16Apr 14, 2026Updated 3 months ago
- [NeurIPS 2025] Official repo of "Martian World Model: Controllable Video Synthesis with Physically Accurate 3D Reconstructions"☆21Aug 6, 2025Updated last year
- [ICCV 2025 Highlight] "Edit360: 2D Image Edits to 3D Assets from Any Angle"☆21Feb 4, 2026Updated 6 months ago
- ViGeo: Towards Consistent Video Geometry Estimation☆130Jul 29, 2026Updated last week
- [Arxiv'25] DINO-Tok: Adapting DINO for Visual Tokenizers☆42Apr 11, 2026Updated 3 months ago
- Official code for paper "How Much 3D Do Video Foundation Models Encode?"☆42Mar 24, 2026Updated 4 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models☆269Jul 23, 2026Updated 2 weeks ago
- Modeling Depth Ambiguity: A Mixture-Density Representation for Flying-Point-Free Depth Estimation☆69Jun 5, 2026Updated 2 months ago
- [ICLR'26] SPRINT: Sparse-Dense Residual Fusion for Efficient Diffusion Transformers☆16Mar 19, 2026Updated 4 months ago
- [ICLR2026] AliTok: Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model☆56Oct 12, 2025Updated 9 months ago
- A simple 3D asset retrieval system based on objaverse. Query any 3D asset using text(CN/EN) or images, inter-modal or cross-modal. Equip…☆16Feb 15, 2026Updated 5 months ago
- [ECCV 2026] Official Implementation of Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction☆18Apr 26, 2026Updated 3 months ago
- [ECCV 2026] Feed-forward, Physically Plausible Human-Scene Reconstruction☆20Jun 24, 2026Updated last month
- [ECCV 2026] PointSplat: Compact Gaussian Splatting via Human-Centric Prediction☆47Jul 14, 2026Updated 3 weeks ago
- [CVPR 2026 Highlight] Clay-to-Stone: Phase-wise 3D Gaussian Splatting for Monocular Articulated Hand-Object Manipulation Modeling☆17Jun 9, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder☆58Jul 24, 2026Updated 2 weeks ago
- [CVPR2026] Exploring Spatial Intelligence from a Generative Perspective☆31Jun 3, 2026Updated 2 months ago
- ☆20Jan 31, 2023Updated 3 years ago
- Video Instance Segmentation with a Propose-Reduce Paradigm (ICCV 2021)☆43Aug 5, 2023Updated 3 years ago
- Official Implementation of ARM4R ICML 2025☆54Sep 18, 2025Updated 10 months ago
- [NeurIPS 2024 Spotlight] CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning.☆14Dec 12, 2024Updated last year
- G3Splat: Geometrically Consistent Generalizable Gaussian Splatting☆68Jan 5, 2026Updated 7 months ago
- ☆35Apr 10, 2026Updated 4 months ago
- View-Invariant Policy Learning via Zero-Shot Novel View Synthesis (CoRL 2024)☆31Sep 28, 2025Updated 10 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Project page for SparseSplat: Towards Applicable Feed-Forward 3D Gaussian Splatting with Pixel-Unaligned Prediction (CVPR 2026)☆33Apr 28, 2026Updated 3 months ago
- Powering 3DGS with Realistic Screen-Space Reflections☆15Apr 7, 2025Updated last year
- A Curated List of Awesome Video World Models with AR Diffusion: Covering Algorithms, Applications, and Infrastructure, Aimed at Serving a…☆686Jun 4, 2026Updated 2 months ago
- World Modeling by Forecasting Vision Foundation Model Features☆52Jul 25, 2026Updated 2 weeks ago
- ☆79Jun 22, 2026Updated last month
- ☆59Jun 8, 2026Updated 2 months ago
- [CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs…☆63Apr 10, 2026Updated 4 months ago