Loc3R-VLM: Language-based Localization and 3D Reasoning with Vision-Language Models
☆49Mar 24, 2026Updated 4 months ago
Alternatives and similar repositories for Loc3R-VLM
Users that are interested in Loc3R-VLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2026] Official implementation of "Splat and Distill: Augmenting Teachers with Feed-Forward 3D Reconstruction For 3D-Aware Distillat…☆29May 4, 2026Updated 3 months ago
- Repository for WACV23 paper "Automatically Annotating Indoor Images with CAD Models via RGB-D Scans"☆18Aug 5, 2025Updated last year
- [CVPR 2026] Official implementation of "GA-VLN: Geometry-Aware BEV Representation for Efficient Vision-Language Navigation"☆16May 25, 2026Updated 2 months ago
- Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)☆21Mar 18, 2026Updated 4 months ago
- ☆19Apr 5, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICCV 2025] Scene Coordinate Reconstruction Priors☆22Nov 10, 2025Updated 8 months ago
- [AAAI 2026] SmartSplat: Feature-Smart Gaussians for Scalable Compression of Ultra-High-Resolution Images☆19Dec 26, 2025Updated 7 months ago
- The code for paper 'Learning from Videos for 3D World: Enhancing MLLMs with 3D Vision Geometry Priors'☆251Nov 28, 2025Updated 8 months ago
- [ICLR'26] This repository is the implementation of "3D Aware Region Prompted Vision Language Model"☆30Feb 19, 2026Updated 5 months ago
- Official code for paper: N3D-VLM: Native 3D Grounding Enables Accurate Spatial Reasoning in Vision-Language Models☆117Jan 14, 2026Updated 6 months ago
- [CVPR2026] Zoo3D: Zero-Shot 3D Object Detection at Scene Level☆36Feb 25, 2026Updated 5 months ago
- [ECCV 2026] Group3D: MLLM-Driven Semantic Grouping for Open-Vocabulary 3D Object Detection☆50Jul 2, 2026Updated last month
- [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"☆151Nov 4, 2025Updated 9 months ago
- [CVPR 2026] Think with 3D: Geometric Imagination Grounded Spatial Reasoning from Limited Views☆245May 7, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A toolkit for computing Video Fréchet Inception Distance (VFID) metrics.☆11May 28, 2024Updated 2 years ago
- From Flatland to Space (SPAR). Accepted to NeurIPS 2025 Datasets & Benchmarks. A large-scale dataset & benchmark for 3D spatial perceptio…☆93Jan 5, 2026Updated 7 months ago
- [CVPR2026] STRNet: Visual Navigation with Spatio-Temporal Representation through Dynamic Graph Aggregation☆18Apr 9, 2026Updated 3 months ago
- WorldCache: Content-Aware Caching for Accelerated Video World Models☆23Jun 28, 2026Updated last month
- [CoRL 2025] ActLoc: Learning to Localize on the Move via Active Viewpoint Selection☆65Sep 21, 2025Updated 10 months ago
- [ICCV 2023] Distilling Coarse-to-fine Semantic Matching Knowledge for Weakly Supervised 3D Visual Grounding☆14Oct 2, 2024Updated last year
- [NeurIPS 2025] Implementation of "SceneSplat++: A Large Dataset and Comprehensive Benchmark for Language Gaussian Splatting"☆56May 16, 2025Updated last year
- ☆35Nov 17, 2025Updated 8 months ago
- ☆47Apr 9, 2026Updated 3 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ECCV 2026🔥] Official code repository for "Stream3D-VLM: Online 3D Spatial Understanding with Incremental Geometry Priors"☆55Jun 23, 2026Updated last month
- ☆30May 13, 2026Updated 2 months ago
- Unifying Visual Localization and Scene Recognition on Panoramic Annular Lens☆13May 18, 2020Updated 6 years ago
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆432Jul 15, 2026Updated 3 weeks ago
- This repo contains VPR models that have been fine-tuned for indoor usage.☆16May 15, 2024Updated 2 years ago
- A Unified Perspective-to-Equirectangular Visual Place Recognition Framework☆21Dec 19, 2025Updated 7 months ago
- ☆16Dec 18, 2025Updated 7 months ago
- Cambrian-S: Towards Spatial Supersensing in Video☆564Apr 3, 2026Updated 4 months ago
- 🔥 Official code repository for "Unlocking Dense Metric Depth Estimation in VLMs"☆155Jul 22, 2026Updated 2 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2026] Official implementation of the paper "📷 On the Generalization Capacities of MLLMs for Spatial Intelligence"☆29Mar 17, 2026Updated 4 months ago
- [CVPR2026] Chain of World: World Model Thinking in Latent Motion☆62Mar 4, 2026Updated 5 months ago
- Source Code for View Consistent Purification for Accurate Cross-View Localization, ICCV 2023☆19Nov 24, 2024Updated last year
- GOTPR: General Outdoor Text-based Place Recognition Using Scene Graph Retrieval with OpenStreetMap☆39Mar 3, 2026Updated 5 months ago
- [NeurIPS 2025] the official project page of a paper, "PLANA3R: Zero-shot Metric Planar 3D Reconstruction via Feed-Forward Planar Splattin…☆75May 4, 2026Updated 3 months ago
- [CVPR'26 Highlight] SimRecon: SimReady Compositional Scene Reconstruction from Real Videos☆137Apr 14, 2026Updated 3 months ago
- ☆56Aug 30, 2024Updated last year