☆20Sep 3, 2025Updated last year
Alternatives and similar repositories for spatial-understanding
Users that are interested in spatial-understanding are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆25Apr 30, 2026Updated 4 months ago
- This is a project on visual spatial reasoning tasks-SIBench☆28Jan 12, 2026Updated 8 months ago
- ☆19Oct 12, 2025Updated 11 months ago
- [ICLR 2026] Official implementation of the paper "📷 On the Generalization Capacities of MLLMs for Spatial Intelligence"☆31Mar 17, 2026Updated 6 months ago
- Scaffold Prompting to promote LMMs☆46Dec 16, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [NeurIPS 2025] Source codes for the paper "MindJourney: Test-Time Scaling with World Models for Spatial Reasoning"☆153Nov 4, 2025Updated 10 months ago
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 7 months ago
- Japanese Converter Kanji to Hiragana, Katakana, Roma-ji☆13Jul 19, 2023Updated 3 years ago
- [NeurIPS 2024 Oral] "Bayesian-Guided Label Mapping for Visual Reprogramming"☆12Dec 20, 2024Updated last year
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs☆71Mar 22, 2026Updated 5 months ago
- [ACL 2023] Are Pre-trained Language Models Useful for Model Ensemble in Chinese Grammatical Error Correction?☆10Dec 15, 2025Updated 9 months ago
- Toward Ambulatory Vision: Learning Visually-Grounded Active View Selection☆25Feb 5, 2026Updated 7 months ago
- super-resolution; post-training quantization; model compression☆14Nov 10, 2023Updated 2 years ago
- [IEEE TMI'23] IOP-FL: Inside-Outside Personalization for Federated Medical Image Segmentation☆13Apr 2, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- SpatialThinker: Reinforcing 3D Reasoning in Multimodal LLMs via Spatial Rewards☆42Jan 28, 2026Updated 7 months ago
- [PR 2024] HTQ: Exploring the High-Dimensional Trade-Off of Mixed-Precision Quantization☆12Jul 16, 2024Updated 2 years ago
- ☆15Mar 21, 2025Updated last year
- ☆13Jan 9, 2024Updated 2 years ago
- Collection of PhD Advice Links☆23Oct 14, 2022Updated 3 years ago
- [ICML 2026] ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning☆90Jul 8, 2026Updated 2 months ago
- [MICCAI2023] Client-Level Differential Privacy via Adaptive Intermediary in Federated Medical Imaging☆17Aug 2, 2024Updated 2 years ago
- Code for CVPR 2023 Robust Generalization against Photon-Limited Corruptions via Worst-Case Sharpness Minimization☆13Mar 27, 2023Updated 3 years ago
- My academic homepage☆15Jan 15, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆41Jan 26, 2023Updated 3 years ago
- t5-model-onnx,中文拼写纠错,Chinese spelling correction。☆15Sep 18, 2022Updated 4 years ago
- [NeurIPS 2025] 3DRS: MLLMs Need 3D-Aware Representation Supervision for Scene Understanding☆160Dec 9, 2025Updated 9 months ago
- Source code for ZePHyR: Zero-shot Pose Hypothesis Rating @ ICRA 2021☆25Aug 17, 2022Updated 4 years ago
- ☆28May 9, 2026Updated 4 months ago
- Code and data for the paper: AI Sees Your Location—But With A Bias Toward The Wealthy World☆19Dec 15, 2025Updated 9 months ago
- A python (PyTorch) implementation of federated multi-encoding U-Net (Fed-MENU) method for federated learning-based multi-organ segmentati…☆16Nov 5, 2024Updated last year
- Code for ThriftyDAgger☆15Dec 29, 2021Updated 4 years ago
- ☆14Mar 1, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- [ICCV 2025] QuantCache:Adaptive Importance-Guided Quantization with Hierarchical Latent and Layer Caching for Video Generation☆18Sep 26, 2025Updated 11 months ago
- The official repo of "Mimir: Hierarchical Goal-Driven Diffusion with Uncertainty Propagation for End-to-End Autonomous Driving"☆21Aug 24, 2026Updated 3 weeks ago
- Implementation of Language-Conditioned Path Planning (Amber Xie, Youngwoon Lee, Pieter Abbeel, Stephen James)☆27Sep 1, 2023Updated 3 years ago
- Official PyTorch implementation of The Linear Attention Resurrection in Vision Transformer☆15Sep 7, 2024Updated 2 years ago
- ☆17Oct 5, 2025Updated 11 months ago
- Pytorch Implementation of Videos as Space-Time Region Graphs☆27Jul 17, 2026Updated 2 months ago
- ☆18Jun 14, 2025Updated last year