[ACM MM'25] Code for the paper "Open3D-VQA: A Benchmark for Embodied Spatial Reasoning with Multimodal Large Language Model in Open Space"
☆18Jul 9, 2026Updated last week
Alternatives and similar repositories for Open3D-VQA.code
Users that are interested in Open3D-VQA.code are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Apr 23, 2026Updated 2 months ago
- [ACL'25 Oral] Code for the paper "UrbanVideo-Bench: Benchmarking Vision-Language Models on Embodied Intelligence with Video Data in Urban…☆31Jul 15, 2025Updated last year
- ☆22Apr 8, 2026Updated 3 months ago
- Benchmark for Multi-robot Cleaning Task Allocation☆14Aug 13, 2023Updated 2 years ago
- [CVPR2024] FCS: Feature Calibration and Separation for Non-Exemplar Class Incremental Learning☆20Apr 18, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A SITL guide for setting up Ardupilot, Gazebo & ROS☆15Jul 27, 2020Updated 5 years ago
- 完全免费的 VPN。亲测有效的科学上网,同时支持 windows、mac、linux、ios 和 andrioid 系统。并提供 chrome、firefox、opera 等浏览器的插件使用。☆11Jul 17, 2018Updated 8 years ago
- [NeurIPS 2024] MSR3D: Multimodal Situated Reasoning in 3D Scenes☆75Dec 2, 2025Updated 7 months ago
- ☆36Apr 18, 2024Updated 2 years ago
- ☆32Jun 24, 2024Updated 2 years ago
- ☆11Jul 12, 2024Updated 2 years ago
- ☆18Jul 11, 2025Updated last year
- ☆65May 13, 2026Updated 2 months ago
- RefDrone: A Challenging Benchmark for Drone Scene Referring Expression Comprehension☆44Jul 8, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆28May 30, 2026Updated last month
- [CVPR 2024] LaMPilot: An Open Benchmark Dataset for Autonomous Driving with Language Model Programs☆43Jan 21, 2026Updated 5 months ago
- Fusion-then-Distillation: Toward Cross-modal Positive Distillation for Domain Adaptive 3D Semantic Segmentation [TCSVT 2025]☆17Feb 21, 2025Updated last year
- [CVPR 2023] Cascade Evidential Learning for Open-world Weakly-supervised Temporal Action Localization☆12Jul 9, 2024Updated 2 years ago
- [ICCVW2025] V-RoAst: Visual Road Assessment. Can VLM be a Road Safety Assessor Using the iRAP Standard?☆13Dec 17, 2025Updated 7 months ago
- Codebase for LangNav paper☆19Jun 13, 2024Updated 2 years ago
- 🔥GrabS in PyTorch (ICLR 2025 Spotlight)☆22May 30, 2026Updated last month
- ☆38Jul 22, 2024Updated last year
- ☆35Feb 26, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Jun 13, 2025Updated last year
- [AAAI2024] An official pytorch implement of the paper: Vision-Language Pre-training with Object Contrastive Learning for 3D Scene Underst…☆13Dec 8, 2024Updated last year
- Official PyTorch implementation of: "Cannot See the Forest for the Trees: Aggregating Multiple Viewpoints to Better Classify Objects in V…☆14Aug 29, 2022Updated 3 years ago
- [ACM MM 2025] ViTCoT: Video-Text Interleaved Chain-of-Thought for Boosting Video Understanding in Large Language Models☆18Jul 15, 2025Updated last year
- 基于InternLm chat 7B大模型基座,构建一个Agent ,可以调用 MMYOLO 工具来完成图像内视觉任务☆11Oct 30, 2024Updated last year
- [CVPR 2025] Official repository of CamPoint : Boosting Point Cloud Segmentation with Virtual Camera☆24Mar 21, 2026Updated 4 months ago
- b站(BILIBILI)抢新年礼物插件☆10Mar 8, 2018Updated 8 years ago
- [JMLR] Gradual Domain Adaptation: Theory and Algorithms☆11Jan 14, 2025Updated last year
- Official Implementation for "SiLVR : A Simple Language-based Video Reasoning Framework"☆19Jan 18, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2024 Highlight] Scaffold-GS: Structured 3D Gaussians for View-Adaptive Rendering☆10Jul 29, 2024Updated last year
- The public reproducible analysis code used for the gaze project☆11May 16, 2026Updated 2 months ago
- 本项目是集成了各大云服务厂商的短信业务平台,支持ThinkPHP5.0、ThinkPHP5.1和ThinkPHP6.0,由宁波晟嘉网络科技有限公司维护,目前支持阿里云、腾讯云、七牛云、又拍云、Ucloud和华为云,您如果有其他厂商的集成需求,请通过邮件联系作者提交需求。☆18Jul 2, 2020Updated 6 years ago
- CVMHT : Complementary-View Multiple Human Tracking (AAAI 2020).☆10Dec 9, 2021Updated 4 years ago
- [IROS'25] COCMT☆12Aug 14, 2025Updated 11 months ago
- This repository contains all the code and data used in our article titled “Estimating international trade status of countries from global…☆10Jul 6, 2023Updated 3 years ago
- Host CIFAR-10.2 Data Set☆13Sep 22, 2021Updated 4 years ago