cshizhe/vil3dref

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/cshizhe/vil3dref)

cshizhe / vil3dref

Official implementation of Language Conditioned Spatial Relation Reasoning for 3D Object Grounding (NeurIPS'22).

☆67

Alternatives and similar repositories for vil3dref

Users that are interested in vil3dref are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

3dlg-hcvc / multi3drefer
View on GitHub
[ICCV 2023] Multi3DRefer: Grounding Text Description to Multiple 3D Objects
☆98Mar 26, 2026Updated 3 months ago
zyang-ur / SAT
View on GitHub
SAT: 2D Semantics Assisted Training for 3D Visual Grounding, ICCV 2021 (Oral)
☆32Sep 29, 2021Updated 4 years ago
rohjunha / language-refer
View on GitHub
☆27Jan 3, 2024Updated 2 years ago
nickgkan / butd_detr
View on GitHub
Code for the ECCV22 paper "Bottom Up Top Down Detection Transformers for Language Grounding in Images and Point Clouds"
☆95Jun 9, 2023Updated 3 years ago
eslambakr / CoT3D_VG
View on GitHub
Chain_of_Thoughts_3D_Visual_Grounding
☆21Apr 20, 2024Updated 2 years ago
Managed Kubernetes at scale on DigitalOcean • Ad
DigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
referit3d / referit3d
View on GitHub
Code accompanying our ECCV-2020 paper on 3D Neural Listeners.
☆141Jun 29, 2021Updated 5 years ago
yanmin-wu / EDA
View on GitHub
[CVPR 2023] EDA: Explicit Text-Decoupling and Dense Alignment for 3D Visual Grounding
☆135Oct 11, 2023Updated 2 years ago
sega-hsj / MVT-3DVG
View on GitHub
[CVPR 2022] Multi-View Transformer for 3D Visual Grounding
☆81Nov 9, 2022Updated 3 years ago
dfki-av / MiKASA-3DVG
View on GitHub
[CVPR'24] MiKASA: Multi-Key-Anchor & Scene-Aware Transformer for 3D Visual Grounding
☆18Dec 13, 2024Updated last year
3d-vista / 3D-VisTA
View on GitHub
Official implementation of ICCV 2023 paper "3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment"
☆215Sep 7, 2023Updated 2 years ago
fjhzhixi / 3D-SPS
View on GitHub
☆64May 17, 2023Updated 3 years ago
joyhsu0504 / NS3D
View on GitHub
☆46Mar 27, 2023Updated 3 years ago
eslambakr / LAR-Look-Around-and-Refer
View on GitHub
This is the official implementation for our paper;"LAR:Look Around and Refer".
☆30Dec 1, 2022Updated 3 years ago
ATR-DBI / ScanQA
View on GitHub
☆161Aug 23, 2023Updated 2 years ago
GPU virtual machines on DigitalOcean Gradient AI • Ad
Get to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
Ivan-Tang-3D / ViewRefer3D
View on GitHub
(ICCV2023) Official implementation of 'ViewRefer: Grasp the Multi-view Knowledge for 3D Visual Grounding with GPT and Prototype Guidance'…
☆60Apr 18, 2024Updated 2 years ago
jianghaojun / Awesome-3D-Vision-and-Language
View on GitHub
A collection of 3D vision and language (e.g., 3D Visual Grounding, 3D Question Answering and 3D Dense Caption) papers and datasets.
☆101Feb 26, 2023Updated 3 years ago
zlccccc / 3DVG-Transformer
View on GitHub
[ICCV2021] 3DVG-Transformer: Relation Modeling for Visual Grounding on Point Clouds
☆43Jul 6, 2022Updated 4 years ago
daveredrum / ScanRefer_Browser
View on GitHub
☆11Feb 1, 2023Updated 3 years ago
CurryYuan / PhraseRefer
View on GitHub
[TNNLS] Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases
☆17Jul 10, 2025Updated last year
SilongYong / SQA3D
View on GitHub
[ICLR 2023] SQA3D for embodied scene understanding and reasoning
☆169Oct 13, 2023Updated 2 years ago
daveredrum / ScanRefer
View on GitHub
[ECCV 2020] ScanRefer: 3D Object Localization in RGB-D Scans using Natural Language
☆303Feb 10, 2023Updated 3 years ago
scene-verse / SceneVerse
View on GitHub
Official implementation of ECCV24 paper "SceneVerse: Scaling 3D Vision-Language Learning for Grounded Scene Understanding"
☆288Mar 19, 2025Updated last year
liudaizong / Awesome-3D-Visual-Grounding
View on GitHub
😎 up-to-date & curated list of awesome 3D Visual Grounding papers, methods & resources.
☆283Jan 14, 2026Updated 6 months ago
Simple, predictable pricing with DigitalOcean hosting • Ad
Always know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
hanhung / TGNN
View on GitHub
☆26Mar 15, 2022Updated 4 years ago
daveredrum / D3Net
View on GitHub
[ECCV2022] D3Net: A Unified Speaker-Listener Architecture for 3D Dense Captioning and Visual Grounding
☆44Aug 27, 2022Updated 3 years ago
leolyj / 3D-VLP
View on GitHub
This is the code related to "Context-aware Alignment and Mutual Masking for 3D-Language Pre-training" (CVPR 2023).
☆29Jun 15, 2023Updated 3 years ago
haomengz / D-LISA
View on GitHub
[NeurIPS‘24] Multi-Object 3D Grounding with Dynamic Modules and Language Informed Spatial Attention
☆28Jun 15, 2025Updated last year
embodied-generalist / embodied-generalist
View on GitHub
[ICML 2024] LEO: An Embodied Generalist Agent in 3D World
☆485Apr 20, 2025Updated last year
vlc-robot / robot_sugar
View on GitHub
Official implementation of "SUGAR: Pre-training 3D Visual Representations for Robotics" (CVPR'24).
☆46Jun 19, 2025Updated last year
PNXD / FFL-3DOG
View on GitHub
Free-form Description-guided 3D Visual Graph Networks for Object Grounding in Point Cloud
☆18Jun 23, 2022Updated 4 years ago
jiemingcui / ProBio
View on GitHub
[NeurIPS'23] "ProBio: A Protocol-guided Multimodal Dataset for Molecular Biology Lab"
☆16Dec 13, 2023Updated 2 years ago
CurryYuan / InstanceRefer
View on GitHub
[ICCV 2021] InstanceRefer: Cooperative Holistic Understanding for Visual Grounding on Point Clouds through Instance Multi-level Contextua…
☆74Mar 22, 2025Updated last year
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
szzexpoi / rex
View on GitHub
Official Repository for CVPR 2022 paper "REX: Reasoning-aware and Grounded Explanation"
☆22Nov 21, 2023Updated 2 years ago
Leon1207 / 3DRefTR
View on GitHub
This is a PyTorch implementation of 3DRefTR proposed by our paper "A Unified Framework for 3D Point Cloud Visual Grounding"
☆26Aug 24, 2023Updated 2 years ago
ZCMax / ScanReason
View on GitHub
[ECCV 2024] Empowering 3D Visual Grounding with Reasoning Capabilities
☆85Oct 10, 2024Updated last year
chunfeng3364 / LARC
View on GitHub
☆19Jun 26, 2024Updated 2 years ago
ZzZZCHS / Chat-Scene
View on GitHub
[NeurIPS 2024 & TPAMI 2026] Chat-Scene: Bridging 3D Scene and Large Language Models with Object Identifiers
☆216Apr 12, 2026Updated 3 months ago
heng-hw / SpaCap3D
View on GitHub
[IJCAI 2022] Spatiality-guided Transformer for 3D Dense Captioning on Point Clouds (official pytorch implementation)
☆21Aug 31, 2022Updated 3 years ago
vlc-robot / hiveformer
View on GitHub
☆33Sep 25, 2024Updated last year