A curated list of awesome papers for reconstructing 4D spatial intelligence from video. (arXiv 2507.21045)
☆516Jun 5, 2026Updated 2 months ago
Alternatives and similar repositories for Awesome-4D-Spatial-Intelligence
Users that are interested in Awesome-4D-Spatial-Intelligence are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆955Oct 27, 2025Updated 9 months ago
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,108Jul 3, 2026Updated last month
- [ICLR 2026] Trace Anything: Representing Any Video in 4D via Trajectory Fields☆558Oct 31, 2025Updated 9 months ago
- [CVPR 2026] Official implementation of "MoVieS: Motion-Aware 4D Dynamic View Synthesis in One Second".☆468Mar 19, 2026Updated 4 months ago
- [ICCV 2025] SpatialTrackerV2: 3D Point Tracking Made Easy☆993Feb 27, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2025 (Spotlight)] The implementation for the paper "4DGT Learning a 4D Gaussian Transformer Using Real-World Monocular Videos"☆474Sep 19, 2025Updated 10 months ago
- Any4D: Unified Feed-Forward Metric 4D Reconstruction☆396Apr 17, 2026Updated 4 months ago
- Official implementation of Continuous 3D Perception Model with Persistent State☆1,479Aug 27, 2025Updated 11 months ago
- [ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling☆490Apr 16, 2026Updated 4 months ago
- Cameras as Relative Positional Encoding☆746Dec 18, 2025Updated 7 months ago
- [CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations☆595Apr 22, 2026Updated 3 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆194Mar 10, 2026Updated 5 months ago
- [ICLR 2026] A simple state update rule to enhance length generalization for CUT3R☆724May 11, 2026Updated 3 months ago
- [ICCV 2025] A simple training-free approach adapting DUSt3R for dynamic scenes.☆534Apr 1, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- MapAnything: Universal Feed-Forward Metric 3D Reconstruction☆3,653Aug 7, 2026Updated last week
- Code for the project "MegaSaM: Accurate, Fast and Robust Structure and Motion from Casual Dynamic Videos"☆1,355Jan 5, 2026Updated 7 months ago
- Code for ICCV'2025 (Best student paper honorable mention) "RayZer: A Self-supervised Large View Synthesis Model"☆446Nov 24, 2025Updated 8 months ago
- Dynamic 3D Foundation Model using Causal Transformer. [ICLR 2026]☆397May 8, 2026Updated 3 months ago
- [ICML 2026] 4RC: 4D Reconstruction via Conditional Querying Anytime and Anywhere☆226Jul 7, 2026Updated last month
- [ICCV 2025 Highlight] Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction☆438Jun 6, 2025Updated last year
- [ICCV 2025 Oral] MVTracker: Multi-view 3D Point Tracking☆515Nov 3, 2025Updated 9 months ago
- ViPE: Video Pose Engine for Geometric 3D Perception☆2,074Jun 9, 2026Updated 2 months ago
- [CVPR 2026] Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction☆371Mar 20, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implement of VGGT-Long☆897Mar 20, 2026Updated 4 months ago
- [ICLR 2025 Oral] Official code for "LVSM: A Large View Synthesis Model with Minimal 3D Inductive Bias"☆551Aug 4, 2025Updated last year
- [ICLR 2026] FastVGGT: Fast Visual Geometry Transformer☆810Jan 28, 2026Updated 6 months ago
- [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling☆608Oct 26, 2025Updated 9 months ago
- [ECCV 2026] Syn4D: A Multiview Synthetic 4D Dataset☆138Jul 4, 2026Updated last month
- [CVPR 2026 Highlight & Best Paper of VideoWorldModel Workshop] NeoVerse: Enhancing 4D World Model with in-the-wild Monocular Videos☆653May 12, 2026Updated 3 months ago
- 🌟A curated list of DUSt3R-related papers and resources, tracking recent advancements using this geometric foundation model.☆803Nov 5, 2025Updated 9 months ago
- TAPIP3D: Tracking Any Point in Persistent 3D Geometry☆420Dec 28, 2025Updated 7 months ago
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆440Jul 15, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [TPAMI 2026] 3D and 4D World Modeling: A Survey☆966Jul 29, 2026Updated 2 weeks ago
- [ICLR2026] Official Implementation of "Dens3R: A Foundation Model for 3D Geometry Prediction"☆395May 14, 2026Updated 3 months ago
- A curated list of awesome 3D scene generation papers. (IJCV 2026)☆1,078Aug 11, 2026Updated last week
- [CVPR 2025] Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video☆227May 25, 2025Updated last year
- ☆722May 1, 2025Updated last year
- Open source impl of **MV-DUSt3R+ Single-Stage Scene Reconstruction from Sparse Views In 2 Seconds** from Meta Reality Labs. Project page …☆603Updated this week
- Stereo4D dataset and processing code☆312Nov 4, 2025Updated 9 months ago