dfki-av/MiKASA-3DVG

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/dfki-av/MiKASA-3DVG)

dfki-av / MiKASA-3DVG

[CVPR'24] MiKASA: Multi-Key-Anchor & Scene-Aware Transformer for 3D Visual Grounding

☆18

Alternatives and similar repositories for MiKASA-3DVG

Users that are interested in MiKASA-3DVG are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

eslambakr / CoT3D_VG
View on GitHub
Chain_of_Thoughts_3D_Visual_Grounding
☆21Apr 20, 2024Updated 2 years ago
cshizhe / vil3dref
View on GitHub
Official implementation of Language Conditioned Spatial Relation Reasoning for 3D Object Grounding (NeurIPS'22).
☆67Dec 2, 2022Updated 3 years ago
dfki-av / sg-pgm
View on GitHub
SG-PGM: Partial Graph Matching Network with Semantic Geometric Fusion for 3D Scene Graph Alignment and Its Downstream Tasks
☆41Jun 10, 2024Updated 2 years ago
tpzou / HGL
View on GitHub
HGL: Hierarchical Geometry Learning for Test-time Adaptation in 3D Point Cloud Segmentation
☆15Sep 13, 2024Updated last year
zlccccc / 3DVG-Transformer
View on GitHub
[ICCV2021] 3DVG-Transformer: Relation Modeling for Visual Grounding on Point Clouds
☆43Jul 6, 2022Updated 4 years ago
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
yanmin-wu / EDA
View on GitHub
[CVPR 2023] EDA: Explicit Text-Decoupling and Dense Alignment for 3D Visual Grounding
☆135Oct 11, 2023Updated 2 years ago
taolinzhang / 3DVLP
View on GitHub
[AAAI2024] An official pytorch implement of the paper: Vision-Language Pre-training with Object Contrastive Learning for 3D Scene Underst…
☆13Dec 8, 2024Updated last year
qzp2018 / MCLN
View on GitHub
This is a PyTorch implementation of MCLN proposed by our paper "Multi-branch Collaborative Learning Network for 3D Visual Grounding"(ECCV…
☆27Oct 10, 2024Updated last year
Ivan-Tang-3D / ViewRefer3D
View on GitHub
(ICCV2023) Official implementation of 'ViewRefer: Grasp the Multi-view Knowledge for 3D Visual Grounding with GPT and Prototype Guidance'…
☆60Apr 18, 2024Updated 2 years ago
ZCMax / ScanReason
View on GitHub
[ECCV 2024] Empowering 3D Visual Grounding with Reasoning Capabilities
☆85Oct 10, 2024Updated last year
joyhsu0504 / NS3D
View on GitHub
☆46Mar 27, 2023Updated 3 years ago
CurryYuan / ZSVG3D
View on GitHub
[CVPR 2024] Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding
☆63Aug 3, 2024Updated last year
TencentARC / TaCA
View on GitHub
Official code for the paper, "TaCA: Upgrading Your Visual Foundation Model with Task-agnostic Compatible Adapter".
☆16Jun 20, 2023Updated 3 years ago
Yioutpi / Awesome-3D-Understanding
View on GitHub
☆13Jul 22, 2024Updated 2 years ago
Managed Kubernetes at scale on DigitalOcean • Ad
DigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
GuanqiaoDing / CNN-CIFAR10
View on GitHub
Implement and Compare VGG, ResNet and ResNeXt on CIFAR-10
☆10Mar 22, 2019Updated 7 years ago
stoneMo / CIGN
View on GitHub
Official implementation for CIGN
☆17Sep 11, 2023Updated 2 years ago
haomengz / D-LISA
View on GitHub
[NeurIPS‘24] Multi-Object 3D Grounding with Dynamic Modules and Language Informed Spatial Attention
☆28Jun 15, 2025Updated last year
zhang-guangyi / HJSCC
View on GitHub
This is a pytorch implementation of our AAAI paper for learned image transmission with HVAE
☆12Mar 2, 2026Updated 4 months ago
OpenQiming / OpenQiming-TrainInference
View on GitHub
The training and inference tools of OpenQiming
☆26May 14, 2026Updated 2 months ago
Lyther / Lecture-Notes
View on GitHub
Lecture notes include almost everything in my notebook.
☆12Sep 8, 2025Updated 10 months ago
csimo005 / SUMMIT
View on GitHub
☆11Oct 4, 2023Updated 2 years ago
yuanzhoulvpi2017 / yuanzhoulvpi2017
View on GitHub
personal info
☆11Mar 23, 2024Updated 2 years ago
AliBahri94 / SVWA_TTA
View on GitHub
[WACV 2025-Oral Presentation] Test-Time Adaptation in Point Clouds: Leveraging Sampling Variation with Weight Averaging
☆13Mar 31, 2025Updated last year
Deploy to Railway using AI coding agents - Free Credits Offer • Ad
Use Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
idejie / 3DSyn
View on GitHub
☆12May 19, 2025Updated last year
leoli646 / Adapter-X
View on GitHub
Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision
☆11Jul 22, 2024Updated 2 years ago
oceanflowlab / QuatRoPE
View on GitHub
[CVPR 2026] Scalable Object Relation Encoding for Better 3D Spatial Reasoning in Large Language Models
☆18May 28, 2026Updated last month
toheart / cocursor
View on GitHub
☆18Feb 9, 2026Updated 5 months ago
Wangdai-0800 / CoordinateAttention_Keras
View on GitHub
A Keras Implementation of Coordinate Attention follows https://github.com/Andrew-Qibin/CoordAttention
☆13Sep 25, 2021Updated 4 years ago
tudelft-iv / Adapting_CVL
View on GitHub
☆12Dec 19, 2024Updated last year
AronCao49 / Latte
View on GitHub
[ECCV 2024] Reliable Spatial-Temporal Voxels for Multi-Modal Test-Time Adaptation
☆18Jan 12, 2026Updated 6 months ago
ajhamdi / vointcloud
View on GitHub
Voint Cloud: Multi-View Point Cloud Representation for 3D Understanding (ICLR 2023)
☆22May 2, 2023Updated 3 years ago
miemie2013 / Keras-DCNv2
View on GitHub
Deformable Convolutional Networks v2 with Keras and Tensorflow1.x
☆19Dec 3, 2020Updated 5 years ago
Wordpress hosting with auto-scaling - Free Trial Offer • Ad
Fully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
ripl / Transcrib3D
View on GitHub
Official Repository of "Transcrib3D: 3D Referring Expression Resolution through Large Language Models" accepted at IROS 2024
☆13Mar 30, 2026Updated 3 months ago
JLUtangchuan / Parts2Words
View on GitHub
This is the source code of Part2Word: Learning Joint Embedding of Point Clouds and Text by Bidirectional Matching between Parts and Words
☆16Mar 22, 2023Updated 3 years ago
ByteDance-Seed / TaskMem
View on GitHub
☆26Jun 2, 2026Updated last month
3dlg-hcvc / multi3drefer
View on GitHub
[ICCV 2023] Multi3DRefer: Grounding Text Description to Multiple 3D Objects
☆98Mar 26, 2026Updated 4 months ago
ShichaoSun / ConAbsSum
View on GitHub
☆13Aug 27, 2021Updated 4 years ago
AIM-SKKU / RA-Touch
View on GitHub
RA-Touch: Retrieval-Augmented Touch Understanding with Enriched Visual Data (ACM MM '25)
☆15Sep 12, 2025Updated 10 months ago
Zhengtq / CoordAttention_tensorflow
View on GitHub
Unofficial implementation of "Coordinate Attention for Efficient Mobile Network Design". CoordAttention tensorflow slim
☆17Mar 10, 2021Updated 5 years ago