Official code of DMA: Dense Multimodal Alignment for Open-Vocabulary 3D Scene Understanding, ECCV 2024
☆32Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for DMA
Users that are interested in DMA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official PyTorch codes for "Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation", ECCV2024☆32Jul 19, 2024Updated 2 years ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆17Aug 1, 2026Updated 2 months ago
- SyncNoise: Geometrically Consistent Noise Prediction for Text-based 3D Scene Editing☆19Dec 28, 2024Updated last year
- Chain_of_Thoughts_3D_Visual_Grounding☆21Apr 20, 2024Updated 2 years ago
- [ECCV2024] ScaleDreamer: Scalable Text-to-3D Synthesis with Asynchronous Score Distillation☆54Mar 28, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2026] DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆27Sep 30, 2026Updated last week
- [NeurIPS 2024] XMask3D: Cross-modal Mask Reasoning for Open Vocabulary 3D Semantic Segmentation☆38Jan 20, 2025Updated last year
- ScatterFormer: Efficient Voxel Transformer with Scattered Linear Attention (ECCV 2024)☆81May 20, 2025Updated last year
- Official code for our Paper "SSL: A Self-similarity Loss for Improving Generative Image Super-resolution" in ACMMM 2024☆51Jun 6, 2026Updated 4 months ago
- [ECCV'24] A novel weakly supervised framework for 3D object detection from 2D bounding boxes. It can easily extend to novel scenarios and…☆36Jul 26, 2024Updated 2 years ago
- Fine-grained Image-to-LiDAR Contrastive Distillation with Visual Foundation Models (NeurIPS2024)☆45Nov 22, 2024Updated last year
- [ICCV 2025] SAS: Segment Any 3D Scene with Integrated 2D Priors☆40Jun 25, 2025Updated last year
- ☆24Jun 29, 2026Updated 3 months ago
- FPR: False Positive Rectification for Weakly Supervised Semantic Segmentation (ICCV 2023)☆24Sep 24, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Project page of "GaussianSR: 3D Gaussian Super-Resolution with 2D Diffusion Priors"☆23Jul 1, 2024Updated 2 years ago
- [ECCV 2024] Empowering 3D Visual Grounding with Reasoning Capabilities☆85Oct 10, 2024Updated 2 years ago
- SimCMF: A Simple Cross-modal Fine-tuning Strategy from Vision Foundation Models to Any Imaging Modality☆34Nov 25, 2024Updated last year
- [NeurIPS24 Spotlight] Voxel Mamba: Group-Free State Space Models for Point Cloud based 3D Object Detection☆163Sep 26, 2024Updated 2 years ago
- Official implementation of the paper "Unifying 3D Vision-Language Understanding via Promptable Queries"☆89Aug 2, 2024Updated 2 years ago
- [ECCV 2024] M3DBench introduces a comprehensive 3D instruction-following dataset with support for interleaved multi-modal prompts.☆61Oct 1, 2024Updated 2 years ago
- ☆12Jul 18, 2024Updated 2 years ago
- ☆99Mar 25, 2024Updated 2 years ago
- [AAAI'26] BEVDilation: LiDAR-Centric Multi-Modal Fusion for 3D Object Detection☆47Dec 3, 2025Updated 10 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆32Apr 29, 2026Updated 5 months ago
- ☆55Oct 3, 2024Updated 2 years ago
- Photo3D: Advancing Photorealistic 3D Generation through Structure‑Aligned Detail Enhancement☆22Mar 18, 2026Updated 6 months ago
- [ICLR 2026] - One2Scene☆54May 25, 2026Updated 4 months ago
- (CVPR 2023) PLA: Language-Driven Open-Vocabulary 3D Scene Understanding & (CVPR2024) RegionPLC: Regional Point-Language Contrastive Learn…☆300Jun 28, 2024Updated 2 years ago
- Open3DIS: Open-vocabulary 3D Instance Segmentation with 2D Mask Guidance (CVPR 2024)☆138Nov 12, 2024Updated last year
- [NeurIPS 2024] Lexicon3D: Probing Visual Foundation Models for Complex 3D Scene Understanding☆102Feb 2, 2025Updated last year
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆182Jul 7, 2025Updated last year
- Code for the paper: "ODIN: A Single Model for 2D and 3D Segmentation" (CVPR 2024)☆178Sep 28, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ECCV2024, LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models☆18Aug 9, 2024Updated 2 years ago
- Code for "Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes"☆59Mar 28, 2024Updated 2 years ago
- [CVPR 2025 Highlight] Official code repository for "Inst3D-LMM: Instance-Aware 3D Scene Understanding with Multi-modal Instruction Tuning…☆134Jan 30, 2026Updated 8 months ago
- [AAAI26] ViP3DE: Fast Multi-view Consistent 3D Editing with Video Priors☆23Aug 9, 2026Updated 2 months ago
- [ECCV2024] Source Prompt Disentangled Inversion for Boosting Image Editability with Diffusion Models☆48Jul 4, 2024Updated 2 years ago
- [ECCV 2026] Official code repository for "Self-transcendence: Is External Feature Guidance Indispensable for Accelerating Diffusion Trans…☆39Aug 25, 2026Updated last month
- Official implementation of CN-RMA: Combined Network with Ray Marching Aggregation for 3D Indoor Object Detection from Multi-view Images☆21Jun 24, 2024Updated 2 years ago