[3DV 2026] Open Vocabulary Monocular 3D Object Detection
☆102Apr 29, 2026Updated 4 months ago
Alternatives and similar repositories for ovmono3d
Users that are interested in ovmono3d are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆52May 6, 2025Updated last year
- [ICCV 2025] Detect Anything 3D in the Wild☆295Dec 14, 2025Updated 9 months ago
- [CVPR 2025] The offical implementation of 'MonoDGP: Monocular 3D Object Detection with Decoupled-Query and Geometry-Error Priors'☆100Aug 13, 2025Updated last year
- [NeurIPS 2025] LabelAny3D: Label Any Object 3D in the Wild☆134Aug 23, 2026Updated 3 weeks ago
- ☆17Jun 29, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ImOV3D: Learning Open Vocabulary Point Clouds 3D Object Detection from Only 2D Images (NeurIPS2024)☆94Feb 20, 2026Updated 7 months ago
- ☆56Jan 2, 2025Updated last year
- [ECCV 2024] LabelDistill: Label-guided Cross-modal Knowledge Distillation for Camera-based 3D Object Detection☆45Nov 1, 2024Updated last year
- A large-scale NOCS dataset.☆101Jul 12, 2024Updated 2 years ago
- MonoDINO-DETR: Depth-Enhanced Monocular 3D Object Detection Using a Vision Foundation Model☆46May 27, 2025Updated last year
- [ICCV'25] 3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection☆123Oct 14, 2025Updated 11 months ago
- Bridging Perspectives: Foundation Model Guided BEV Maps for 3D Object Detection and Tracking☆21Oct 13, 2025Updated 11 months ago
- Official code for paper: N3D-VLM: Native 3D Grounding Enables Accurate Spatial Reasoning in Vision-Language Models☆119Jan 14, 2026Updated 8 months ago
- [ AAAI 2026 ] The official implementation of 'MonoCLUE: Object-Aware Clustering Enhances Monocular 3D Object Detection'☆22Mar 23, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 3D BBox refinement interface used in LabelAny3D (NeurIPS 2025)☆25Aug 22, 2026Updated 3 weeks ago
- Code release for our NeurIPS 2023 paper "Uni3DETR: Unified 3D Detection Transformer", our ECCV 2024 paper "OV-Uni3DETR: Towards Unified O…☆121Jul 29, 2024Updated 2 years ago
- ☆99Mar 25, 2024Updated 2 years ago
- Vision-Language Guidance for LiDAR-based Unsupervised 3D Object Detection☆29Nov 21, 2024Updated last year
- ☆14Oct 6, 2024Updated last year
- [CVPR 2026 Highlight] MonoCoP: Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection☆26Aug 3, 2026Updated last month
- Code release for "Omni3D A Large Benchmark and Model for 3D Object Detection in the Wild"☆855Apr 7, 2024Updated 2 years ago
- [ECCV2024] Weakly Supervised 3D Object Detection via Multi-Level Visual Guidance☆23Jul 14, 2024Updated 2 years ago
- [ICLR 2025 (Oral 📢) ] Our OpenYOLO3D model achieves state-of-the-art performance in Open Vocabulary 3D Instance Segmentation on ScanNet2…☆261Mar 17, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official code for NeurIPS2023 paper CoDA: Collaborative Novel Box Discovery and Cross-modal Alignment for Open-vocabulary 3D Object Detec…☆224May 28, 2026Updated 3 months ago
- [ECCV 2024] RecurrentBEV: A Long-term Temporal Fusion Framework for Multi-view 3D Detection☆34Sep 28, 2024Updated last year
- [ICCV 2025] Boosting Multi-View Indoor 3D Object Detection via Adaptive 3D Volume Construction☆29Oct 1, 2025Updated 11 months ago
- [ECCV 2024] Omni6DPose: A Benchmark and Model for Universal 6D Object Pose Estimation and Tracking☆139Sep 1, 2024Updated 2 years ago
- [NeurIPS 2025] OpenAD: Open-World Autonomous Driving Benchmark for 3D Object Detection☆72Nov 28, 2024Updated last year
- [ECCV'24] Approaching Outside: Scaling Unsupervised 3D Object Detection from 2D Scene.☆40Sep 3, 2024Updated 2 years ago
- Spatial Aptitude Training for Multimodal Langauge Models☆34Feb 8, 2026Updated 7 months ago
- Code release for the ECCV 2024 paper 'Fully Test-Time Adaptation for Monocular 3D Object Detection'☆58Dec 10, 2024Updated last year
- [ICCV 2023] GeoMIM: towards better 3d knowledge transfer via masked image modeling for multi-view 3d understanding☆53Aug 28, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Unsupervised 3D Object Detection [NeurIPS 2024]☆45Feb 12, 2026Updated 7 months ago
- ☆17Jun 21, 2026Updated 2 months ago
- Code for the Boxer research paper☆657Jul 28, 2026Updated last month
- [CVPR'25] SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding☆224Apr 21, 2025Updated last year
- This repository contains the code for the paper - "Aligning Text, Images, and 3D Structure Token-by-Token" (CVPR 2026)☆49Jun 11, 2025Updated last year
- ImageNet3D: Towards General-Purpose Object-Level 3D Understanding☆22Dec 6, 2024Updated last year
- [NeurIPS 2023] 3D Copy-Paste: Physically Plausible Object Insertion for Monocular 3D Detection☆58Mar 27, 2024Updated 2 years ago