[NeurIPS 2025 Spotlight] Official implementation of the SIU3R: Simultaneous Scene Understanding and 3D Reconstruction Beyond Feature Alignment
☆164Sep 25, 2025Updated 10 months ago
Alternatives and similar repositories for SIU3R
Users that are interested in SIU3R are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR'26] IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction☆427Dec 1, 2025Updated 7 months ago
- Reasoning in Space via Grounding in the World (ICLR 2025)☆56Nov 3, 2025Updated 8 months ago
- [NeurIPS 2025] Streaming 3D Reconstruction with Explicit Spatial Pointer Memory☆192Mar 10, 2026Updated 4 months ago
- [CVPR 2026 Highlight] Unified 3D Reconstruction and Semantic Understanding via Generalizable Gaussian Splatting from Unposed Multi-View I…☆196Apr 10, 2026Updated 3 months ago
- [NeurIPS'24] Large Spatial Model: End-to-end Unposed Images to Semantic 3D☆236Feb 11, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2026] A simple state update rule to enhance length generalization for CUT3R☆710May 11, 2026Updated 2 months ago
- Code for ICCV'2025 (Best student paper honorable mention) "RayZer: A Self-supervised Large View Synthesis Model"☆444Nov 24, 2025Updated 8 months ago
- Official implementation of “4D LangVGGT: 4D Language-Visual Geometry Grounded Transformer”☆90Mar 25, 2026Updated 4 months ago
- [CVPR25] SPC-GS: Gaussian Splatting with Semantic-Prompt Consistency for Indoor Open-World Free-view Synthesis from Sparse Inputs☆20Aug 27, 2025Updated 10 months ago
- [3DV 2026 Oral] Official Repo of "SAIL-Recon: Large SfM by Augmenting Scene Regression with Localization"☆299Feb 23, 2026Updated 5 months ago
- [ICLR2026] Official Implementation of "Dens3R: A Foundation Model for 3D Geometry Prediction"☆395May 14, 2026Updated 2 months ago
- [ECCV 2026] Walking in the Implicit: Interactive World Exploration via Neural Scene Representation☆48Jun 30, 2026Updated 3 weeks ago
- PanSt3R: Multi-view Consistent Panoptic Segmentation (official code)☆80Mar 20, 2026Updated 4 months ago
- [ICLR 2026] π^3: Permutation-Equivariant Visual Geometry Learning☆2,088Jul 3, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICCV 2025] LangScene-X: Reconstruct Generalizable 3D Language-Embedded Scenes with TriMap Video Diffusion☆302Jul 15, 2025Updated last year
- ☆35Nov 17, 2025Updated 8 months ago
- [ICML2025 Oral] ReferSplat: Referring Segmentation in 3D Gaussian Splatting☆146May 26, 2026Updated 2 months ago
- [ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling☆604Oct 26, 2025Updated 9 months ago
- [CVPR'26] PE3R: Perception-Efficient 3D Reconstruction. Take 2 - 3 photos with your phone, upload them, wait a few minutes, and then star…☆414Feb 28, 2026Updated 4 months ago
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆346Apr 18, 2026Updated 3 months ago
- [ICCV 2025 Oral] SceneSplat - Gaussian Splatting-based Scene Understanding with Vision-Language Pretraining☆354May 25, 2026Updated 2 months ago
- [CVPR 2026] Gen3R: 3D Scene Generation Meets Feed-Forward Reconstruction☆366Mar 20, 2026Updated 4 months ago
- ☆51Jul 8, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of “4D LangSplat: 4D Language Gaussian Splatting via Multimodal Large Language Models” (CVPR 2025)☆205Oct 10, 2025Updated 9 months ago
- ☆129Jun 17, 2025Updated last year
- [CVPR 2026] This repository is the official implementation of MVGGT: Multimodal Visual Geometry Grounded Transformer for Multiview 3D Ref…☆128Mar 24, 2026Updated 4 months ago
- [CVPR 2026] "E-RayZer: Self-supervised 3D Reconstruction as Spatial Visual Pre-training" official implementation.☆301May 30, 2026Updated last month
- Official implement of VGGT-Long☆884Mar 20, 2026Updated 4 months ago
- [SIGGRAPH Asia 2025 (ACM TOG)] AnySplat: Feed-forward 3D Gaussian Splatting from Unconstrained Views☆897Dec 22, 2025Updated 7 months ago
- [Preprint] Any 3D Scene is Worth 1K Tokens: 3D-Grounded Representation for Scene Generation at Scale☆56Apr 14, 2026Updated 3 months ago
- [NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence☆480Feb 5, 2026Updated 5 months ago
- [NeurIPS 2025] LangSplatV2: High-dimensional 3D Language Gaussian Splatting with 450+ FPS☆243Oct 17, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICCV 2025] A simple training-free approach adapting DUSt3R for dynamic scenes.☆532Apr 1, 2025Updated last year
- [ICLR 2026] Streaming 4D Visual Geometry Transformer☆943Oct 27, 2025Updated 8 months ago
- ☆213Oct 22, 2025Updated 9 months ago
- [ICLR 2026 Oral (top 1.2%)] Official implementation of DepthLM☆363Jun 1, 2026Updated last month
- UniGeo: Taming Video Diffusion for Unified Consistent Geometry Estimation☆136Jun 10, 2025Updated last year
- [ICLR 2024] USB-NeRF: Unrolling Shutter Bundle Adjusted Neural Radiance Fields☆14Mar 24, 2024Updated 2 years ago
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆430Jul 15, 2026Updated last week