☆27Jun 5, 2025Updated last year
Alternatives and similar repositories for 3DLLM-Mem
Users that are interested in 3DLLM-Mem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Neural network methods for multimodal map reconstruction and their usage for robot navigation and control☆15Jun 11, 2024Updated 2 years ago
- [CVPR 2025] Source codes for the paper "3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning"☆271Oct 2, 2025Updated 10 months ago
- ☆74Feb 12, 2026Updated 6 months ago
- ☆74Mar 29, 2025Updated last year
- [ICCV'25] Ross3D: Reconstructive Visual Instruction Tuning with 3D-Awareness☆71Jul 22, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2024 & TPAMI 2026] Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers☆217Apr 12, 2026Updated 4 months ago
- ☆267Aug 6, 2025Updated last year
- [CVPR 2025] The code for paper ''Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding''.☆221Jun 4, 2025Updated last year
- [CVPR 2025] 3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer☆101May 26, 2025Updated last year
- [NeurIPS 2025] 3DRS: MLLMs Need 3D-Aware Representation Supervision for Scene Understanding☆160Dec 9, 2025Updated 8 months ago
- [ICCV 2025] Improving 3D Large Language Model via Robust Instruction Tuning☆71Oct 19, 2025Updated 10 months ago
- [ICLR 2026] UniUGG: Unified 3D Understanding and Generation via Geometric-Semantic Encoding☆63Jul 16, 2026Updated last month
- [CVPR 2026] LightSplat: Fast and Memory-Efficient Open-Vocabulary 3D Scene Understanding in Five Seconds☆29Aug 6, 2026Updated 3 weeks ago
- [CVPR 2024] "LL3DA: Visual Interactive Instruction Tuning for Omni-3D Understanding, Reasoning, and Planning"; an interactive Large Langu…☆319Jul 17, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- SpaceR: The first MLLM empowered by SG-RLVR for video spatial reasoning☆111Jul 9, 2025Updated last year
- Code&Data for Grounded 3D-LLM with Referent Tokens☆136Jan 5, 2025Updated last year
- Spa3R: Predictive Spatial Field Modeling for 3D Visual Reasoning☆52Mar 25, 2026Updated 5 months ago
- ☆19Aug 7, 2025Updated last year
- ☆55Oct 3, 2024Updated last year
- ☆42Jul 16, 2025Updated last year
- [NeurIPS 2025] Official implementation of "RoboRefer: Towards Spatial Referring with Reasoning in Vision-Language Models for Robotics"☆266Dec 16, 2025Updated 8 months ago
- [CVPR 2025] Beacon3D: Object-centric Evaluation for 3D Grounding-QA☆28Nov 25, 2025Updated 9 months ago
- [ECCV 2024] Empowering 3D Visual Grounding with Reasoning Capabilities☆85Oct 10, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆56Jan 9, 2025Updated last year
- ☆25Apr 30, 2026Updated 4 months ago
- LogiCity@NeurIPS'24, D&B track. A multi-agent inductive learning environment for "abstractions".☆27Jun 10, 2025Updated last year
- The code for paper 'Learning from Videos for 3D World: Enhancing MLLMs with 3D Vision Geometry Priors'☆256Nov 28, 2025Updated 9 months ago
- MemoryEQA☆28May 4, 2026Updated 3 months ago
- [ICLR 26] pySpatial: Generating 3D Visual Programs for Zero-Shot Spatial Reasoning☆32Jun 2, 2026Updated 2 months ago
- This is the official evaluation code for Robobench☆24Aug 16, 2026Updated 2 weeks ago
- Collection of the latest spatial, 3D, and video/temporal reasoning papers☆37Sep 29, 2025Updated 11 months ago
- ScaleMaster-Dataset☆17May 11, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆40May 29, 2025Updated last year
- [ECCV 2026] Spatial-TTT: Streaming Visual-based Spatial Intelligence with Test-Time Training☆260Jun 19, 2026Updated 2 months ago
- [CVPR 2026] G2VLM: Geometry Grounded Vision Language Model with Unified 3D Reconstruction and Spatial Reasoning☆352Apr 18, 2026Updated 4 months ago
- [ICCV 2025] Controllable 3D Outdoor Scene Generation via Scene Graphs☆71Jul 13, 2026Updated last month
- OVSegDT, a lightweight transformer policy to solve Open-vocabulary Object Goal Navigation☆19May 25, 2026Updated 3 months ago
- Fast Spatial Memory with Elastic Test-Time Training (4D-LRM + 4D-LVSM)☆106Jun 20, 2026Updated 2 months ago
- Embodied Question Answering (EQA) benchmark and method (ICCV 2025)☆61Aug 12, 2025Updated last year