☆32Jun 24, 2024Updated 2 years ago
Alternatives and similar repositories for EmbSpatial-Bench
Users that are interested in EmbSpatial-Bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Jun 19, 2026Updated 3 months ago
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 10 months ago
- ☆152Jun 17, 2026Updated 3 months ago
- ☆15Feb 24, 2023Updated 3 years ago
- ☆10Nov 16, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This is the official PyTorch implementation of the CVPR 2023 paper: "GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot A…☆10Mar 17, 2024Updated 2 years ago
- ☆12Jul 16, 2024Updated 2 years ago
- A Vision-Language Model for Spatial Affordance Prediction in Robotics☆231Jul 17, 2025Updated last year
- [ECCV 2024] Official implementation of C-Instructor: Controllable Navigation Instruction Generation with Chain of Thought Prompting☆30Dec 16, 2024Updated last year
- ☆17Mar 24, 2026Updated 5 months ago
- DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments☆25Apr 8, 2025Updated last year
- [IROS 24] Official repository of "Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation". We…☆20Apr 1, 2026Updated 5 months ago
- ☆24May 28, 2025Updated last year
- Official implementation of Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions (Ne…☆58Dec 20, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code of the paper "Correctable Landmark Discovery via Large Models for Vision-Language Navigation" (TPAMI 2024)☆16Jun 7, 2024Updated 2 years ago
- Official implementation of "SUGAR: Pre-training 3D Visual Representations for Robotics" (CVPR'24).☆46Jun 19, 2025Updated last year
- [NeurIPS'24] This repository is the implementation of "SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models"☆338Dec 14, 2024Updated last year
- ☆24Oct 8, 2023Updated 2 years ago
- [ICLR 2025 Oral] Official Implementation for "Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Un…☆23Oct 24, 2024Updated last year
- ☆74Jan 8, 2025Updated last year
- [NeurIPS 2025] Official implementation of "RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics"☆266Dec 16, 2025Updated 9 months ago
- Official repository for GraphEQA☆32Sep 25, 2025Updated 11 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 7 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- SimX-OR: Extending Any Simulation Benchmark to Evaluate the Observational Robustness of VLA Models☆32Nov 4, 2025Updated 10 months ago
- [AAAI'26] Official implementation of CMMCoT: Enhancing Complex Multi-Image Comprehension via Multi-Modal Chain-of-Thought and Memory Augm…☆11Dec 5, 2025Updated 9 months ago
- Official implementation of "Grounded Entity-Landmark Adaptive Pre-training for Vision-and-Language Navigation" (ICCV 2023 Oral)☆21Oct 21, 2023Updated 2 years ago
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation (ICCV 2025)☆26Jan 30, 2026Updated 7 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 7 months ago
- Planning as In-Painting: A Diffusion-Based Embodied Task Planning Framework for Environments under Uncertainty☆23Dec 11, 2023Updated 2 years ago
- ☆14Apr 24, 2023Updated 3 years ago
- OpenEQA Embodied Question Answering in the Era of Foundation Models☆364Sep 20, 2024Updated last year
- [ICCV 2025] Official code for Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation☆66Sep 12, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2026] Official code for PairFlow: Closed-Form Source-Target Coupling for Few-Step Generation in Discrete Flow Models☆18Jul 3, 2026Updated 2 months ago
- SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery☆15Feb 1, 2026Updated 7 months ago
- Embodied Question Answering (EQA) benchmark and method (ICCV 2025)☆61Aug 12, 2025Updated last year
- [CVPR 2026 Highlight] Official code for MatLat: Material Latent Space for PBR Texture Generation☆23Jul 16, 2026Updated 2 months ago
- Open Vocabulary Object Navigation☆144May 15, 2025Updated last year
- ☆24Mar 9, 2023Updated 3 years ago
- ☆64Dec 14, 2024Updated last year