[ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence
☆21Jan 26, 2026Updated 8 months ago
Alternatives and similar repositories for SpaCE-10
Users that are interested in SpaCE-10 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SpaceDG: Benchmarking Spatial Intelligence under Visual Degradation☆31Jul 9, 2026Updated 2 months ago
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆29Sep 24, 2026Updated 2 weeks ago
- ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the st…☆25Aug 9, 2026Updated last month
- [TCSVT2026] The official repository of the paper "Contending with Depth Ambiguity in Monocular 3D Pose Estimation via Multi-Hypothesis Mo…☆13Updated this week
- Official repository for Interleave-VLA☆23Apr 5, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces☆86Jun 6, 2025Updated last year
- Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports☆70Mar 15, 2026Updated 6 months ago
- ICLR 2022 paper☆16May 6, 2022Updated 4 years ago
- [ICLR 2026] OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Models☆95Jan 21, 2026Updated 8 months ago
- [NeurIPS 2026] PhotoFlow: Agentic 3D Virtual Photography Missions☆43Sep 26, 2026Updated last week
- (ECCV2024) Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence☆20Jun 27, 2025Updated last year
- ☆17May 24, 2023Updated 3 years ago
- The official repository of the paper "SGA-INTERACT: A 3D Skeleton-based Benchmark for Group Activity Understanding in Modern Basketball T…☆18Mar 11, 2025Updated last year
- [ICLR 2026] Official repository for "Real-3DQA"☆55Apr 1, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [AAAI 26] Official PyTorch implementation of Earth-Adapter: Bridge the Geospatial Domain Gaps with Mixture of Frequency Adaptation☆65May 29, 2025Updated last year
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 6 months ago
- [IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models.☆93Feb 12, 2026Updated 7 months ago
- Official Implementation of Our ICLR 2023 paper "ROCO: A General Framework for Evaluating Robustness of Combinatorial Optimization Solvers…☆20Oct 23, 2024Updated last year
- [CVPR 2024] Situational Awareness Matters in 3D Vision Language Reasoning☆44Dec 9, 2024Updated last year
- [Remote Sensing 2026] Co-Training Vision Language Models for Remote Sensing Multi-task Learning☆41Jul 8, 2026Updated 3 months ago
- [ECCV2024] The official repository of the paper "Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estima…☆18Nov 21, 2024Updated last year
- [NeurIPS 2024 Datasets and Benchmarks Track] Benchmarking PtO and PnO Methods in the Predictive Combinatorial Optimization Regime☆26Mar 27, 2025Updated last year
- Official repository for "TrustGeoGen: Formal-Verified Data Engine for Trustworthy Multi-modal Geometric Problem Solving"☆23Sep 1, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The official repo of CrossEarth-SAR, a sar-centric and billion-scale geospatial foundation model for cross-domain semantic segmentation☆48Mar 18, 2026Updated 6 months ago
- Official PyTorch implementation of ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder☆27Aug 1, 2026Updated 2 months ago
- Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)☆23Mar 18, 2026Updated 6 months ago
- R3-Avatar: Record and Retrieve Temporal Codebook for Reconstructing Photorealistic Human Avatars☆23Nov 23, 2025Updated 10 months ago
- [ECCV2026] Visual Spatial Tuning☆212Mar 25, 2026Updated 6 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆23Apr 24, 2026Updated 5 months ago
- [TGRS'25] AirSpatialBot: A Spatially-Aware Aerial Agent for Fine-Grained Vehicle Attribute Recognization and Retrieval☆33Jan 6, 2026Updated 9 months ago
- [ECCV'26] Code repo for "EvoTok: A Unified Image Tokenizer via Residual Latent Evolution for Visual Understanding and Generation"☆23Jun 18, 2026Updated 3 months ago
- Learning 1D Causal Visual Representation with De-focus Attention Networks☆35Jun 7, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2026] GenExam: A Multidisciplinary Text-to-Image Exam☆72Sep 26, 2026Updated last week
- [ISPRS2026] DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models☆33Mar 24, 2026Updated 6 months ago
- [NeurIPS 2024] MSR3D: Multimodal Situated Reasoning in 3D Scenes☆78Dec 2, 2025Updated 10 months ago
- ☆40May 20, 2025Updated last year
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆310May 14, 2026Updated 4 months ago
- [ICLR'25] Official repo of "PointOBB-v2: Towards Simpler, Faster, and Stronger Single Point Supervised Oriented Object Detection"☆38Mar 27, 2025Updated last year
- [AAAI2025] Multi-clue Consistency Learning to Bridge Gaps Between General and Oriented Object in Semi-supervised Detection☆34Jun 26, 2025Updated last year