Evaluation script for RoboSpatial-Home, a benchmark for spatial reasoning in 2D and 3D vision-language models.
☆23May 14, 2026Updated 3 months ago
Alternatives and similar repositories for RoboSpatial-Eval
Users that are interested in RoboSpatial-Eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆38Aug 25, 2025Updated last year
- Official implementation of the paper "Perturb-and-Revise: Flexible 3D Editing with Generative Trajectories" (CVPR`25)☆15Apr 27, 2026Updated 4 months ago
- [CVPR25 Highlight] Official implementation of Fun3DU, a method for functional understanding and segmentation in 3D scenes☆52Sep 30, 2025Updated 11 months ago
- ☆26Mar 26, 2025Updated last year
- The implementation and supplementary material for our RA-L work "An Affordance Keypoint Detection Network for Robot Manipulation".☆32Jun 15, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of "HiERO: understanding the hierarchy of human behavior enhances reasoning on egocentric videos", accepted at IC…☆17May 22, 2026Updated 3 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- ☆19Jun 22, 2025Updated last year
- [RSS 2025] CLIP-RT : Learning Language-Conditioned Robotic Policies from Natural Language Supervision☆37May 13, 2025Updated last year
- Language Segment-Anything (with updated dependencies)☆37Mar 4, 2024Updated 2 years ago
- An official implementation of RoDyGS: Robust Dynamic Gaussian Splatting for Casual Videos☆36Dec 11, 2024Updated last year
- ☆41Aug 27, 2024Updated 2 years ago
- A real-world autonomous driving simulator based on 3D Gaussian Splatting for scene augmentation☆16Jun 10, 2024Updated 2 years ago
- [ICML 2025] Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation☆52Feb 3, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆17Jan 21, 2026Updated 7 months ago
- A PyTorch implementation of BCO☆12Jun 19, 2023Updated 3 years ago
- Code accompanying paper "SharpDepth: Sharpening Metric Depth Predictions Using Diffusion Distillation"☆29May 8, 2026Updated 3 months ago
- ☆11Aug 7, 2026Updated 3 weeks ago
- Localization via embodied dialog on the navigation graph☆15Apr 18, 2022Updated 4 years ago
- ☆11Nov 30, 2023Updated 2 years ago
- ☆16Jan 13, 2023Updated 3 years ago
- ☆18Feb 13, 2025Updated last year
- ☆37Jun 13, 2026Updated 2 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Getting Started in Imitation Learning☆13Mar 3, 2025Updated last year
- ☆14Dec 5, 2025Updated 8 months ago
- [ICCV 2025] RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping☆50Nov 21, 2025Updated 9 months ago
- 🤗 LeRobot: Making AI for Robotics more accessible with end-to-end learning☆27Oct 18, 2025Updated 10 months ago
- Code for paper on ICRA 2022 workshop on Deformable Object Manipulation. In this work we learn keypoints from synthetic data for robotic c…☆15Aug 6, 2024Updated 2 years ago
- [ICML 2025] OTTER: A Vision-Language-Action Model with Text-Aware Visual Feature Extraction☆119Apr 14, 2025Updated last year
- ☆47Jan 16, 2024Updated 2 years ago
- ☆71Nov 7, 2024Updated last year
- (3DV 2026 Oral) L4P -- a feed-forward foundational model designed for multiple low-level 4D vision perception tasks.☆75Dec 9, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Repository with the source code and models for the paper "Do we still need ImageNet pre-training in remote sensing scene classification?"☆15Jun 20, 2022Updated 4 years ago
- Codebase for the paper: "TIM: A Time Interval Machine for Audio-Visual Action Recognition"☆54Nov 7, 2024Updated last year
- List of papers on video-centric robot learning☆23Nov 16, 2024Updated last year
- A Vision-Language Model for Spatial Affordance Prediction in Robotics☆229Jul 17, 2025Updated last year
- ☆14Jan 5, 2022Updated 4 years ago
- ☆11Jul 26, 2024Updated 2 years ago
- ☆24Feb 10, 2026Updated 6 months ago