[NeurIPS 2026] SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning
☆402Sep 14, 2026Updated 2 weeks ago
Alternatives and similar repositories for SpatialClaw
Users that are interested in SpatialClaw are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Mar 24, 2026Updated 6 months ago
- S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence☆93Jul 22, 2026Updated 2 months ago
- ☆134Aug 27, 2026Updated last month
- code release☆44Jun 22, 2026Updated 3 months ago
- [ICML 2026 Oral] Holi-Spatial: Evolving Video Streams into Holistic 3D Spatial Intelligence☆391Jul 26, 2026Updated 2 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆23Mar 31, 2026Updated 6 months ago
- [ICML 2026] ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning☆90Jul 8, 2026Updated 2 months ago
- Official implementation of "AnthroTAP: Learning Point Tracking with Real-World Motion"☆34Jun 23, 2026Updated 3 months ago
- Official Implementation of "Geometrically-Constrained Agent for Spatial Reasoning"☆93Apr 7, 2026Updated 5 months ago
- ☆17Jun 19, 2026Updated 3 months ago
- [CVPR'26 Highlight] SimRecon: SimReady Compositional Scene Reconstruction from Real Videos☆146Apr 14, 2026Updated 5 months ago
- [ICLR 26] pySpatial: Generating 3D Visual Programs for Zero-Shot Spatial Reasoning☆33Jun 2, 2026Updated 4 months ago
- View planning with multi-turn VLM agents: ViewSuite 6-DoF benchmark on real ScanNet scenes + iterative RL-SFT training☆25Sep 6, 2026Updated 3 weeks ago
- Official code for "Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models"☆59Aug 20, 2026Updated last month
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR 2026] MMSI-Bench: A Benchmark for Multi-Image Spatial Intelligence☆109Apr 28, 2026Updated 5 months ago
- Describe Anything, Anywhere, at Any Moment (DAAAM), a novel approach to real-time, large-scale, spatio-temporal memory☆528Sep 21, 2026Updated last week
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆310May 14, 2026Updated 4 months ago
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆357Apr 18, 2026Updated 5 months ago
- [NeurIPS 2026] Official implementation of paper "VLM³: Vision Language Models Are Native 3D Learners".☆436Sep 24, 2026Updated last week
- From Flatland to Space (SPAR). Accepted to NeurIPS 2025 Datasets & Benchmarks. A large-scale dataset & benchmark for 3D spatial perceptio…☆97Jan 5, 2026Updated 8 months ago
- ☆36Aug 12, 2026Updated last month
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 7 months ago
- STI-Bench : Are MLLMs Ready for Precise Spatial-Temporal World Understanding?☆40Jan 12, 2026Updated 8 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environments☆85Apr 16, 2026Updated 5 months ago
- [NeurIPS 2025] Reinforcing Spatial Reasoning in Vision-Language Models with Interwoven Thinking and Visual Drawing☆100Jul 27, 2025Updated last year
- MMSI-Video-Bench: A Holistic Benchmark for Video-Based Spatial Intelligence☆62Mar 11, 2026Updated 6 months ago
- Cambrian-S: Towards Spatial Supersensing in Video☆570Apr 3, 2026Updated 6 months ago
- Official implementation of "Geometric Action Model for Robot Policy Learning"☆200Sep 21, 2026Updated last week
- ☆26Mar 26, 2025Updated last year
- The official repo for SpaceVista: All-Scale Visual Spatial Reasoning from mm to km.☆44May 26, 2026Updated 4 months ago
- [NeurIPS 2025 Spotlight] Official implementation of Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence☆491Feb 5, 2026Updated 7 months ago
- [CVPR 2026] VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction☆449Jul 15, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [CVPR'26] Official implementation of "C3G: Learning Compact 3D Representations with 2K Gaussians"☆197Jun 2, 2026Updated 4 months ago
- Official implementation of "InterRVOS: Interaction-aware Referring Video Object Segmentation".☆35May 1, 2026Updated 5 months ago
- ☆37Apr 13, 2026Updated 5 months ago
- [ECCV 2026] Real-Time Interactive Multi-Target Video Segmentation☆70Jul 10, 2026Updated 2 months ago
- Qwen-AgentWorld: Language World Models for General Agents☆1,021Jul 20, 2026Updated 2 months ago
- [ICLR 2026 Oral (top 1.2%)] Official implementation of DepthLM☆369Jun 1, 2026Updated 4 months ago
- Official Codebase for "DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos" (ICML 2026)☆1,108Mar 21, 2026Updated 6 months ago