☆21Apr 17, 2025Updated last year
Alternatives and similar repositories for UniM-OV3D
Users that are interested in UniM-OV3D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository for "Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models"☆22Dec 2, 2025Updated 9 months ago
- Official implementation of ICCV 2025 paper "TACO: Taming Diffusion for in-the-wild Video Amodal Completion"☆31Jul 4, 2025Updated last year
- (CVPR 2023) PLA: Language-Driven Open-Vocabulary 3D Scene Understanding & (CVPR2024) RegionPLC: Regional Point-Language Contrastive Learn…☆300Jun 28, 2024Updated 2 years ago
- A Simple Active-and-Adaptive Baseline for Cross-Domain 3D Semantic Segmentation☆13Dec 22, 2022Updated 3 years ago
- [AAAI 2023 Oral] Language-Assisted 3D Feature Learning for Semantic Scene Understanding☆12Aug 1, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆187Jul 5, 2023Updated 3 years ago
- IoU of 2D / 3D rotated bounding box by Pytorch☆13Sep 7, 2021Updated 5 years ago
- TokenAR: Multiple Subject Generation via Autoregressive Token-level enhancement☆21Aug 4, 2026Updated last month
- Repo of "MsSVT: Mixed-scale Sparse Voxel Transformer for 3D Object Detection on Point Clouds".☆37Sep 20, 2023Updated 2 years ago
- ☆23May 26, 2025Updated last year
- Self-supervised adversarial masking for point clouds☆11Jul 12, 2023Updated 3 years ago
- [IJCAI2024] Implementation of "DCDet: Dynamic Cross-based 3D Object Detector"☆16Aug 28, 2024Updated 2 years ago
- PoinTramba: A Hybrid Transformer-Mamba Framework for Point Cloud Analysis☆72May 17, 2025Updated last year
- Open-Vocabulary SAM3D: Understand Any 3D Scene☆44Jun 9, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Self-Supervised Learning for Fine-Grained Image Categorization☆26Dec 18, 2022Updated 3 years ago
- [MM 2023] Toward High Quality Facial Representation Learning☆19Oct 30, 2023Updated 2 years ago
- project page of GaussianFluent☆22Jul 3, 2026Updated 2 months ago
- Graph attention network for stroke classification☆17Jul 25, 2024Updated 2 years ago
- SAMPro3D: Locating SAM Prompts in 3D for Zero-Shot Instance Segmentation (3DV 2025)☆171Apr 17, 2025Updated last year
- [TNNLS] Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases☆17Jul 10, 2025Updated last year
- The official implementation of the paper DBQ-SSD: Dynamic Ball Query for Efficient 3D Object Detection (ICLR 2023)☆18Sep 17, 2023Updated 2 years ago
- Implement "Novel algorithms for 3D surface point cloud boundary detection and edge reconstruction" using Python☆28Mar 11, 2025Updated last year
- Official implementation of MTM☆21Aug 30, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- OpenScan: A Benchmark for Generalized Open-Vocabulary 3D Scene Understanding☆24Jul 6, 2026Updated 2 months ago
- The code for "VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by VIdeo SpatioTemporal Augmentation" [CVPR2025]☆20Feb 27, 2025Updated last year
- ☆99Mar 25, 2024Updated 2 years ago
- BGPSeg: Boundary-Guided Primitive Instance Segmentation of Point Clouds☆28Apr 28, 2025Updated last year
- ☆75Mar 23, 2026Updated 5 months ago
- [CVPR 2024] Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding☆64Aug 3, 2024Updated 2 years ago
- Code release for our NeurIPS 2023 paper "Uni3DETR: Unified 3D Detection Transformer", our ECCV 2024 paper "OV-Uni3DETR: Towards Unified O…☆121Jul 29, 2024Updated 2 years ago
- [CVPR 2025] Beacon3D: Object-centric Evaluation for 3D Grounding-QA☆28Nov 25, 2025Updated 9 months ago
- VCapsBench: A Large-scale Fine-grained Benchmark for Video Caption Quality Evaluation☆20Jun 2, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [Preprint 2022] “Can We Solve 3D Vision Tasks Starting from A 2D Vision Transformer?” by Yi Wang, Zhiwen Fan, Tianlong Chen, Hehe Fan, Zh…☆63Jan 18, 2023Updated 3 years ago
- Codes for "Efficient Multi-Modal 3D Object Detector via Instance Level Contrastive Distillation"☆33Jun 24, 2025Updated last year
- [ECCV2022] Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection☆25Jul 22, 2022Updated 4 years ago
- An object detection codebase based on MegEngine.☆28Dec 14, 2022Updated 3 years ago
- Mask3D predicts accurate 3D semantic instances achieving state-of-the-art on ScanNet, ScanNet200, S3DIS and STPLS3D.☆745Oct 29, 2023Updated 2 years ago
- [ICLR 2024] AGILE3D: Attention Guided Interactive Multi-object 3D Segmentation☆130Apr 1, 2026Updated 5 months ago
- [CVPR 2023 Highlight] Masked Image Modeling with Local Multi-Scale Reconstruction☆57Jul 10, 2023Updated 3 years ago