[ICLR 2026] SpaCE-10: A Comprehensive Benchmark for Multimodal Large Language Models in Compositional Spatial Intelligence
☆20Jan 26, 2026Updated 7 months ago
Alternatives and similar repositories for SpaCE-10
Users that are interested in SpaCE-10 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ECCV'26] GRADE: Grounded Reasoning Assessment for Discipline-informed Editing☆29Aug 16, 2026Updated 2 weeks ago
- ⚽️🤖 Benchmarking LLMs and deep-research agents on real-world football prediction — from the tactical "who scores in minute 67" to the st…☆24Aug 9, 2026Updated 3 weeks ago
- The official repository of the paper "X as Supervision: Contending with Depth Ambiguity in Unsupervised Monocular 3D Pose Estimation"☆13Jan 22, 2025Updated last year
- Official repository for Interleave-VLA☆22Apr 5, 2026Updated 4 months ago
- Stepping VLMs onto the Court: Benchmarking Spatial Intelligence in Sports☆71Mar 15, 2026Updated 5 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ICLR 2022 paper☆16May 6, 2022Updated 4 years ago
- PhotoFlow: Agentic 3D Virtual Photography Missions☆41May 27, 2026Updated 3 months ago
- (ECCV2024) Within the Dynamic Context: Inertia-aware 3D Human Modeling with Pose Sequence☆20Jun 27, 2025Updated last year
- ☆17May 24, 2023Updated 3 years ago
- The official repository of the paper "SGA-INTERACT: A 3D Skeleton-based Benchmark for Group Activity Understanding in Modern Basketball T…☆18Mar 11, 2025Updated last year
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 5 months ago
- [IGARSS 2025 Oral] A Simple Aerial Detection Baseline of Multimodal Language Models.☆92Feb 12, 2026Updated 6 months ago
- Official Implementation of Our ICLR 2023 paper "ROCO: A General Framework for Evaluating Robustness of Combinatorial Optimization Solvers…☆20Oct 23, 2024Updated last year
- [CVPR 2024] Situational Awareness Matters in 3D Vision Language Reasoning☆44Dec 9, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆20Aug 14, 2024Updated 2 years ago
- [Remote Sensing 2026] Co-Training Vision Language Models for Remote Sensing Multi-task Learning☆38Jul 8, 2026Updated last month
- [ECCV2024] The official repository of the paper "Mask as Supervision: Leveraging Unified Mask Information for Unsupervised 3D Pose Estima…☆18Nov 21, 2024Updated last year
- [NeurIPS 2024 Datasets and Benchmarks Track] Benchmarking PtO and PnO Methods in the Predictive Combinatorial Optimization Regime☆26Mar 27, 2025Updated last year
- Official repository for "TrustGeoGen: Formal-Verified Data Engine for Trustworthy Multi-modal Geometric Problem Solving"☆23Sep 1, 2025Updated 11 months ago
- The official repo of CrossEarth-SAR, a sar-centric and billion-scale geospatial foundation model for cross-domain semantic segmentation☆46Mar 18, 2026Updated 5 months ago
- Official PyTorch implementation of ProCLIP: Progressive Vision-Language Alignment via LLM-based Embedder☆26Aug 1, 2026Updated 3 weeks ago
- Open studio for "Thinking with Spatial Code" (https://arxiv.org/pdf/2603.05591)☆21Mar 18, 2026Updated 5 months ago
- R3-Avatar: Record and Retrieve Temporal Codebook for Reconstructing Photorealistic Human Avatars☆23Nov 23, 2025Updated 9 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The open-sourced code of the paper 《Compress Any Segment Anything Model (SAM)》☆19May 29, 2026Updated 3 months ago
- [ECCV2026] Visual Spatial Tuning☆206Mar 25, 2026Updated 5 months ago
- Bridging the gap between image generation and real-world design: a benchmark for structured, multi-constraint commercial visual content g…☆22Apr 24, 2026Updated 4 months ago
- [TGRS'25] AirSpatialBot: A Spatially-Aware Aerial Agent for Fine-Grained Vehicle Attribute Recognization and Retrieval☆33Jan 6, 2026Updated 7 months ago
- [ECCV'26] Code repo for "EvoTok: A Unified Image Tokenizer via Residual Latent Evolution for Visual Understanding and Generation"☆22Jun 18, 2026Updated 2 months ago
- [ICML 2026] GenExam: A Multidisciplinary Text-to-Image Exam☆70May 26, 2026Updated 3 months ago
- [ISPRS2026] DVGBench: Implicit-to-Explicit Visual Grounding Benchmark in UAV Imagery with Large Vision-Language Models☆33Mar 24, 2026Updated 5 months ago
- [NeurIPS 2024] MSR3D: Multimodal Situated Reasoning in 3D Scenes☆77Dec 2, 2025Updated 8 months ago
- ☆40May 20, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models☆297May 14, 2026Updated 3 months ago
- [ICLR'25] Official repo of "PointOBB-v2: Towards Simpler, Faster, and Stronger Single Point Supervised Oriented Object Detection"☆38Mar 27, 2025Updated last year
- [AAAI2025] Multi-clue Consistency Learning to Bridge Gaps Between General and Oriented Object in Semi-supervised Detection☆34Jun 26, 2025Updated last year
- Official repository for the FIRM Reward series☆48Updated this week
- [ICCV'25] Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness☆71Jul 22, 2025Updated last year
- [IJCV] PointOBB-v3: Expanding Performance Boundaries of Single Point-Supervised Oriented Object Detection☆43Sep 25, 2025Updated 11 months ago
- [NeurIPS 2024] CLOVER: Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation☆136Sep 8, 2025Updated 11 months ago