[ICLR 2025 Oral] Official Implementation for "Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Under Ambiguities"
☆23Oct 24, 2024Updated last year
Alternatives and similar repositories for COMFORT
Users that are interested in COMFORT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Concept Learning Dynamics☆17Oct 29, 2024Updated last year
- [ICCV 2025] Official code for Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation☆66Sep 12, 2025Updated 10 months ago
- Can 3D Vision-Language Models Truly Understand Natural Language?☆20Mar 28, 2024Updated 2 years ago
- [ICCV 2023] Going Beyond Nouns With Vision & Language Models Using Synthetic Data☆13Sep 30, 2023Updated 2 years ago
- SpatialThinker: Reinforcing 3D Reasoning in Multimodal LLMs via Spatial Rewards☆42Jan 28, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [NeurIPS'24] This repository is the implementation of "SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models"☆337Dec 14, 2024Updated last year
- [ICCV 2025] PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes☆65Oct 3, 2025Updated 10 months ago
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆22Nov 18, 2025Updated 8 months ago
- [CVPR 2026 Fingdings] This repo is the official implementation of "Euclid’s Gift: Enhancing Spatial Perception and Reasoning in Vision‑La…☆28Mar 15, 2026Updated 4 months ago
- ☆20Mar 2, 2026Updated 5 months ago
- [CVPR 2025] Program synthesis for 3D spatial reasoning☆63Jun 16, 2025Updated last year
- ☆19Mar 28, 2024Updated 2 years ago
- Spatial Aptitude Training for Multimodal Langauge Models☆33Feb 8, 2026Updated 6 months ago
- Official repository for the ICCV2023 paper SAFE: Sensitivity-Aware Features for Out-of-Distribution Object Detection☆14Jul 28, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICCV'25] Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness☆70Jul 22, 2025Updated last year
- ☆12Jan 10, 2025Updated last year
- [NeurIPS'25] SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning☆40Oct 14, 2025Updated 9 months ago
- ☆14Jun 22, 2023Updated 3 years ago
- [ECCV 2026 Oral] Official implementation of "Make Geometry Matter for Spatial Reasoning"☆54Updated this week
- Code and dataset for paper "SpatialBench: Benchmarking Multimodal Large Language Models for Spatial Cognition"☆19Mar 17, 2026Updated 4 months ago
- Code for the paper "If at First You Don't Succeed, Try, Try Again: Faithful Diffusion-based Text-to-Image Generation by Selection"☆27Jul 10, 2023Updated 3 years ago
- ☆32Jun 24, 2024Updated 2 years ago
- ☆16Oct 12, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 这个仓库包含了我在上人工智能课时完成的拼音输入法作业。☆11Feb 16, 2022Updated 4 years ago
- Official Code for the NeurIPS'23 paper "3D-Aware Visual Question Answering about Parts, Poses and Occlusions"☆21Oct 17, 2024Updated last year
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs☆71Mar 22, 2026Updated 4 months ago
- ☆16Jan 30, 2025Updated last year
- ☆19Oct 12, 2025Updated 9 months ago
- [ICLR'26] This repository is the implementation of "3D Aware Region Prompted Vision Language Model"☆30Feb 19, 2026Updated 5 months ago
- Source code for the Paper "Mind the Gap: Benchmarking Spatial Reasoning in Vision-Language Models"☆19Feb 1, 2026Updated 6 months ago
- Code for "Understanding-in-Generation:Reinforcing Generative Capability of Unified Model via Infusing Understanding into Generation"☆15Nov 11, 2025Updated 9 months ago
- Code and data setup for the paper "Are Diffusion Models Vision-and-language Reasoners?"☆33Mar 15, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Training recipe for SpatialReasoner [NeurIPS 2025]☆45Apr 5, 2026Updated 4 months ago
- [ICLR 2026] Official implementation of the paper "📷 On the Generalization Capacities of MLLMs for Spatial Intelligence"☆29Mar 17, 2026Updated 4 months ago
- Official Repo for TiPToP: A Modular Open-Vocabulary Planning System for Robotic Manipulation☆143Updated this week
- [ICRA 2025] RACER: Rich Language-Guided Failure Recovery Policies for Imitation Learning☆47Oct 10, 2024Updated last year
- A paper list for spatial reasoning☆771Jan 19, 2026Updated 6 months ago
- Official PyTorch implementation of the paper ‘CLIP-GS: CLIP-Informed Gaussian Splatting for Real-time and View-consistent 3D Semantic Und…☆58Apr 25, 2024Updated 2 years ago
- "Comp4D: Compositional 4D Scene Generation", Dejia Xu*, Hanwen Liang*, Neel P. Bhatt, Hezhen Hu, Hanxue Liang, Konstantinos N. Platanioti…☆78Apr 27, 2026Updated 3 months ago