[ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence
☆20Jan 26, 2026Updated 5 months ago
Alternatives and similar repositories for SpaCE-10
Users that are interested in SpaCE-10 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation☆31Jul 9, 2026Updated last week
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆28Apr 23, 2026Updated 2 months ago
- ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the st…☆18Updated this week
- The official repository of the paper "X as Supervision: Contending with Depth Ambiguity in Unsupervised Monocular 3D Pose Estimation"☆13Jan 22, 2025Updated last year
- Official repository for Interleave-VLA☆21Apr 5, 2026Updated 3 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces☆86Jun 6, 2025Updated last year
- Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports☆70Mar 15, 2026Updated 4 months ago
- [ICLR 2026] OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models☆88Jan 21, 2026Updated 5 months ago
- PhotoFlow: Agentic 3D Virtual Photography Missions☆38May 27, 2026Updated last month
- (ECCV2024) Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence☆20Jun 27, 2025Updated last year
- ☆17May 24, 2023Updated 3 years ago
- The official repository of the paper "SGA-INTERACT: A 3D Skeleton-based Benchmark for Group Activity Understanding in Modern Basketball T…☆18Mar 11, 2025Updated last year
- [AAAI 26] Official PyTorch implementation of Earth-Adapter: Bridge the Geospatial Domain Gaps with Mixture of Frequency Adaptation☆63May 29, 2025Updated last year
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models.☆92Feb 12, 2026Updated 5 months ago
- Official Implementation of Our ICLR 2023 paper "ROCO: A General Framework for Evaluating Robustness of Combinatorial Optimization Solvers…☆20Oct 23, 2024Updated last year
- [CVPR 2024] Situational Awareness Matters in 3D Vision Language Reasoning☆44Dec 9, 2024Updated last year
- ☆19Aug 14, 2024Updated last year
- [Remote Sensing 2026] Co-Training Vision Language Models for Remote Sensing Multi-task Learning☆37Jul 8, 2026Updated last week
- [ECCV2024] The official repository of the paper "Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estima…☆18Nov 21, 2024Updated last year
- [NeurIPS 2024 Datasets and Benchmarks Track] Benchmarking PtO and PnO Methods in the Predictive Combinatorial Optimization Regime☆26Mar 27, 2025Updated last year
- Official repository for "TrustGeoGen: Formal-Verified Data Engine for Trustworthy Multi-modal Geometric Problem Solving"☆23Sep 1, 2025Updated 10 months ago
- The official repo of CrossEarth-SAR, a sar-centric and billion-scale geospatial foundation model for cross-domain semantic segmentation☆46Mar 18, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official PyTorch implementation of ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder☆25Dec 4, 2025Updated 7 months ago
- Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)☆20Mar 18, 2026Updated 4 months ago
- The open-sourced code of the paper 《Compress Any Segment Anything Model (SAM)》☆19May 29, 2026Updated last month
- R3-Avatar: Record and Retrieve Temporal Codebook for Reconstructing Photorealistic Human Avatars☆23Nov 23, 2025Updated 7 months ago
- [ECCV2026] Visual Spatial Tuning☆198Mar 25, 2026Updated 3 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆20Apr 24, 2026Updated 2 months ago
- [ECCV 2024] Official Implementation of CoPT: Unsupervised Domain Adaptive Segmentation using Domain-Agnostic Text Embeddings☆10Feb 24, 2025Updated last year
- [TGRS'25] AirSpatialBot: A Spatially-Aware Aerial Agent for Fine-Grained Vehicle Attribute Recognization and Retrieval☆32Jan 6, 2026Updated 6 months ago
- [ECCV'26] Code repo for "EvoTok: A Unified Image Tokenizer via Residual Latent Evolution for Visual Understanding and Generation"☆22Jun 18, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Learning 1D Causal Visual Representation with De-focus Attention Networks☆35Jun 7, 2024Updated 2 years ago
- [ISPRS2026] DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models☆30Mar 24, 2026Updated 3 months ago
- [NeurIPS 2024] MSR3D: Multimodal Situated Reasoning in 3D Scenes☆75Dec 2, 2025Updated 7 months ago
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆289May 14, 2026Updated 2 months ago
- KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation☆22Apr 23, 2025Updated last year
- Code for 'Answer Matching Outperforms Multiple Choice for Language Model Evaluation' paper☆18Jul 4, 2025Updated last year
- [ICLR'25] Official repo of "PointOBB-v2: Towards Simpler, Faster, and Stronger Single Point Supervised Oriented Object Detection"☆38Mar 27, 2025Updated last year