Official code of DMA: Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding, ECCV 2024
☆32Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for DMA
Users that are interested in DMA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official PyTorch codes for "Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation", ECCV2024☆31Jul 19, 2024Updated 2 years ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆16Aug 1, 2026Updated 3 weeks ago
- SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing☆19Dec 28, 2024Updated last year
- Chain_of_Thoughts_3D_Visual_Grounding☆21Apr 20, 2024Updated 2 years ago
- [ECCV2024] ScaleDreamer: Scalable Text-to-3D Synthesis with Asynchronous Score Distillation☆53Mar 28, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆26Jun 13, 2026Updated 2 months ago
- [NeurIPS 2024] XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation☆37Jan 20, 2025Updated last year
- ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention (ECCV 2024)☆80May 20, 2025Updated last year
- Official code for our Paper "SSL: A Self-similarity Loss for Improving Generative Image Super-resolution" in ACMMM 2024☆51Jun 6, 2026Updated 2 months ago
- [ECCV'24] A novel weakly supervised framework for 3D object detection from 2D bounding boxes. It can easily extend to novel scenarios and…☆36Jul 26, 2024Updated 2 years ago
- Fine-grained Image-to-LiDAR Contrastive Distillation with Visual Foundation Models (NeurIPS2024)☆45Nov 22, 2024Updated last year
- [ICCV 2025] SAS: Segment Any 3D Scene with Integrated 2D Priors☆40Jun 25, 2025Updated last year
- ☆23Jun 29, 2026Updated 2 months ago
- FPR: False Positive Rectification for Weakly Supervised Semantic Segmentation (ICCV 2023)☆24Sep 24, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Project page of "GaussianSR: 3D Gaussian Super-Resolution with 2D Diffusion Priors"☆23Jul 1, 2024Updated 2 years ago
- [ECCV 2024] Empowering 3D Visual Grounding with Reasoning Capabilities☆85Oct 10, 2024Updated last year
- SimCMF: A Simple Cross-modal Fine-tuning Strategy from Vision Foundation Models to Any Imaging Modality☆34Nov 25, 2024Updated last year
- [NeurIPS24 Spotlight] Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection☆164Sep 26, 2024Updated last year
- Official implementation of the paper "Unifying 3D Vision-Language Understanding via Promptable Queries"☆86Aug 2, 2024Updated 2 years ago
- [ECCV 2024] M3DBench introduces a comprehensive 3D instruction-following dataset with support for interleaved multi-modal prompts.☆61Oct 1, 2024Updated last year
- ☆12Jul 18, 2024Updated 2 years ago
- ☆99Mar 25, 2024Updated 2 years ago
- [AAAI'26] BEVDilation: LiDAR-Centric Multi-Modal Fusion for 3D Object Detection☆42Dec 3, 2025Updated 8 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆32Apr 29, 2026Updated 4 months ago
- ☆55Oct 3, 2024Updated last year
- Photo3D: Advancing Photorealistic 3D Generation through Structure‑Aligned Detail Enhancement☆22Mar 18, 2026Updated 5 months ago
- [ICLR 2026] - One2Scene☆51May 25, 2026Updated 3 months ago
- (CVPR 2023) PLA: Language-Driven Open-Vocabulary 3D Scene Understanding & (CVPR2024) RegionPLC: Regional Point-Language Contrastive Learn…☆300Jun 28, 2024Updated 2 years ago
- Open3DIS: Open-vocabulary 3D Instance Segmentation with 2D Mask Guidance (CVPR 2024)☆138Nov 12, 2024Updated last year
- [NeurIPS 2024] Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding☆102Feb 2, 2025Updated last year
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆179Jul 7, 2025Updated last year
- Code for the paper: "ODIN: A Single Model for 2D and 3D Segmentation" (CVPR 2024)☆178Feb 27, 2026Updated 6 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ECCV2024, LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models☆18Aug 9, 2024Updated 2 years ago
- [CVPR 2025 Highlight] Official code repository for "Inst3D-LMM: Instance-Aware 3D Scene Understanding with Multi-modal Instruction Tuning…☆135Jan 30, 2026Updated 7 months ago
- Code for "Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes"☆59Mar 28, 2024Updated 2 years ago
- [AAAI26] ViP3DE: Fast Multi-view Consistent 3D Editing with Video Priors☆22Aug 9, 2026Updated 3 weeks ago
- [ECCV 2026] Official code repository for "Self-transcendence: Is External Feature Guidance Indispensable for Accelerating Diffusion Trans…☆38Updated this week
- [ECCV2024] Source Prompt Disentangled Inversion for Boosting Image Editability with Diffusion Models☆48Jul 4, 2024Updated 2 years ago
- Official implementation of CN-RMA: Combined Network with Ray Marching Aggregation for 3D Indoor Object Detection from Multi-view Images☆21Jun 24, 2024Updated 2 years ago