☆32Jun 24, 2024Updated 2 years ago
Alternatives and similar repositories for EmbSpatial-Bench
Users that are interested in EmbSpatial-Bench are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Jun 19, 2026Updated 2 months ago
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 9 months ago
- ☆150Jun 17, 2026Updated 2 months ago
- ☆15Feb 24, 2023Updated 3 years ago
- This is the official PyTorch implementation of the CVPR 2023 paper: "GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot A…☆10Mar 17, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Jul 16, 2024Updated 2 years ago
- A Vision-Language Model for Spatial Affordance Prediction in Robotics☆229Jul 17, 2025Updated last year
- [ECCV 2024] Official implementation of C-Instructor: Controllable Navigation Instruction Generation with Chain of Thought Prompting☆31Dec 16, 2024Updated last year
- ☆17Mar 24, 2026Updated 5 months ago
- DOZE: A Dataset for Open-Vocabulary Zero-Shot Object Navigation in Dynamic Environments☆25Apr 8, 2025Updated last year
- [IROS 24] Official repository of "Mind the Error! Detection and Localization of Instruction Errors in Vision-and-Language Navigation". We…☆19Apr 1, 2026Updated 4 months ago
- ☆24May 28, 2025Updated last year
- Official implementation of Human-Aware Vision-and-Language Navigation: Bridging Simulation to Reality with Dynamic Human Interactions (Ne…☆58Dec 20, 2024Updated last year
- Code of the paper "Correctable Landmark Discovery via Large Models for Vision-Language Navigation" (TPAMI 2024)☆16Jun 7, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of "SUGAR: Pre-training 3D Visual Representations for Robotics" (CVPR'24).☆46Jun 19, 2025Updated last year
- [NeurIPS'24] This repository is the implementation of "SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models"☆337Dec 14, 2024Updated last year
- ☆24Oct 8, 2023Updated 2 years ago
- [ICLR 2025 Oral] Official Implementation for "Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Un…☆23Oct 24, 2024Updated last year
- ☆75Jan 8, 2025Updated last year
- [NeurIPS 2025] Official implementation of "RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics"☆266Dec 16, 2025Updated 8 months ago
- Official repository for GraphEQA☆32Sep 25, 2025Updated 11 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆20Feb 1, 2026Updated 6 months ago
- [AAAI'26] Official implementation of CMMCoT: Enhancing Complex Multi-Image Comprehension via Multi-Modal Chain-of-Thought and Memory Augm…☆11Dec 5, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- SimX-OR: Extending Any Simulation Benchmark to Evaluate the Observational Robustness of VLA Models☆34Nov 4, 2025Updated 9 months ago
- Official implementation of "Grounded Entity-Landmark Adaptive Pre-training for Vision-and-Language Navigation" (ICCV 2023 Oral)☆21Oct 21, 2023Updated 2 years ago
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation (ICCV 2025)☆26Jan 30, 2026Updated 7 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- ☆14Apr 24, 2023Updated 3 years ago
- Planning as In-Painting: A Diffusion-Based Embodied Task Planning Framework for Environments under Uncertainty☆23Dec 11, 2023Updated 2 years ago
- OpenEQA Embodied Question Answering in the Era of Foundation Models☆365Sep 20, 2024Updated last year
- [ICCV 2025] Official code for Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation☆66Sep 12, 2025Updated 11 months ago
- [ICLR 2026] Official code for PairFlow: Closed-Form Source-Target Coupling for Few-Step Generation in Discrete Flow Models☆18Jul 3, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Embodied Question Answering (EQA) benchmark and method (ICCV 2025)☆61Aug 12, 2025Updated last year
- SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery☆15Feb 1, 2026Updated 6 months ago
- [CVPR 2026 Highlight] Official code for MatLat: Material Latent Space for PBR Texture Generation☆21Jul 16, 2026Updated last month
- Open Vocabulary Object Navigation☆143May 15, 2025Updated last year
- ☆25Mar 9, 2023Updated 3 years ago
- ☆64Dec 14, 2024Updated last year
- ☆47May 13, 2024Updated 2 years ago