Collection of the latest spatial, 3D, and video/temporal reasoning papers
☆37Sep 29, 2025Updated 11 months ago
Alternatives and similar repositories for awesome-spatial-reasoning
Users that are interested in awesome-spatial-reasoning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Awesome Spatial Intelligence (Personal Use)☆58Jan 7, 2026Updated 7 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- SpatialDreamer: Incentivizing Spatial Reasoning via Active Mental Imagery☆15Feb 1, 2026Updated 6 months ago
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆19Jun 2, 2026Updated 2 months ago
- ☆71Feb 4, 2026Updated 6 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A paper list for spatial reasoning☆778Aug 23, 2026Updated last week
- Github repository for "Why Is Spatial Reasoning Hard for VLMs? An Attention Mechanism Perspective on Focus Areas" (ICML 2025)☆75May 2, 2025Updated last year
- ☆27Jun 5, 2025Updated last year
- Official code for the paper "Adversarial Magnification to Deceive Deepfake Detection through Super Resolution"☆12Jun 26, 2023Updated 3 years ago
- ☆74Feb 12, 2026Updated 6 months ago
- SpaceR: The first MLLM empowered by SG-RLVR for video spatial reasoning☆111Jul 9, 2025Updated last year
- Official repo for From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models☆34Nov 2, 2025Updated 9 months ago
- Benchmark and training code for MindCube: spatial mental modeling in vision-language models from limited views.☆168Aug 23, 2026Updated last week
- [ICLR 2026] Uni-CoT: Towards Unified Chain-of-Thought Reasoning Across Text and Vision☆236May 31, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [Electronics'21] Facial Emotion Recognition Using Transfer Learning in the Deep CNN"☆22Jul 3, 2024Updated 2 years ago
- [CVPR 2026]SpatialStack: Layered Geometry-Language Fusion for 3D VLM Spatial Reasoning☆35Aug 4, 2026Updated 3 weeks ago
- Official Implementation of "CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning" on MIC…☆18Feb 12, 2025Updated last year
- ☆19Apr 10, 2025Updated last year
- Localization via embodied dialog on the navigation graph☆15Apr 18, 2022Updated 4 years ago
- [WACV 2026]Official Code of the paper “Equivariant Sampling for Improving Diffusion Model-based Image Restoration“☆19Jan 29, 2026Updated 7 months ago
- Package for the teleoperation of UR5+Allegro Robot composite☆12Apr 28, 2022Updated 4 years ago
- 一体化网页笔记批注、协作与专注辅助工具。 All-in-one Chrome extension for annotated learning, real-time collaboration, and reading focus tools.☆19Updated this week
- This repository collects and organises state‑of‑the‑art papers on spatial reasoning for Multimodal Vision–Language Models (MVLMs).☆322Feb 17, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆10Nov 18, 2021Updated 4 years ago
- ☆18Jul 31, 2025Updated last year
- A Visualization Tool for GPU Occupancy on S Cluster.☆13Nov 16, 2022Updated 3 years ago
- SimX-OR: Extending Any Simulation Benchmark to Evaluate the Observational Robustness of VLA Models☆34Nov 4, 2025Updated 9 months ago
- Official repository for “Reasoning in the Dark: Interleaved Vision-Text Reasoning in Latent Space”☆18Jan 27, 2026Updated 7 months ago
- Official pytorch implementation of "RITUAL: Random Image Transformations as a Universal Anti-hallucination Lever in Large Vision Language…☆14Dec 16, 2024Updated last year
- [NeurIPS DB 2025] A gym environment for visual spatial reasoning - knot so simple :)☆19Jun 9, 2026Updated 2 months ago
- Code for paper 'MulViMotion: Shape-aware 3D Myocardial Motion Tracking from Multi-View Cardiac MRI'☆13Sep 2, 2022Updated 3 years ago
- [KDD 2024] Team up GBDTs and DNNs: Advancing Efficient and Effective Tabular Prediction with Tree-hybrid MLPs☆13Mar 3, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS'24] This repository is the implementation of "SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models"☆337Dec 14, 2024Updated last year
- Companion repository to "Prompt Compression and Contrastive Conditioning for Controllability and Toxicity Reduction in Language Models"☆14May 31, 2023Updated 3 years ago
- Envision: Benchmarking Unified Understanding & Generation for Causal World Process Insights☆32Jan 9, 2026Updated 7 months ago
- Coda and Data for NeurIPS 2025 paper "MuSLR: Multimodal Symbolic Logical Reasoning"☆17Oct 5, 2025Updated 10 months ago
- Code for the ECCV22 paper Demystifying Unsupervised Semantic Correspondence Estimation☆14Oct 18, 2022Updated 3 years ago
- [ECCV 2026 Oral] Official implementation of "Make Geometry Matter for Spatial Reasoning"☆55Aug 6, 2026Updated 3 weeks ago
- [ICCV 2025] Official PyTorch Code for "Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval"☆19Aug 23, 2025Updated last year