princeton-vl/Rel3D

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/princeton-vl/Rel3D)

princeton-vl / Rel3D

Official code for NeurRIPS 2020 paper "Rel3D: A Minimally Contrastive Benchmark for Grounding Spatial Relations in 3D"

☆33

Alternatives and similar repositories for Rel3D

Users that are interested in Rel3D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

princeton-vl / think_visually
View on GitHub
Code for ACL 2018 paper 'Think Visually: Question Answering through Virtual Imagery'
☆13Mar 24, 2023Updated 3 years ago
AlvinWen428 / spatial-relation-benchmark
View on GitHub
☆15Oct 12, 2024Updated last year
Sina-Baharlou / Depth-VRD
View on GitHub
Improving Visual Relation Detection using Depth Maps (ICPR 2020)
☆47Jul 24, 2022Updated 4 years ago
uvavision / SyViC
View on GitHub
[ICCV 2023] Going Beyond Nouns With Vision & Language Models Using Synthetic Data
☆13Sep 30, 2023Updated 2 years ago
daveredrum / ScanRefer_Browser
View on GitHub
☆11Feb 1, 2023Updated 3 years ago
AI Agents on DigitalOcean Gradient AI Platform • Ad
Build production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
referit3d / referit3d
View on GitHub
Code accompanying our ECCV-2020 paper on 3D Neural Listeners.
☆141Jun 29, 2021Updated 5 years ago
jpeyre / analogy
View on GitHub
Code for the paper "Detecting visual relations using analogies", ICCV19
☆22Jan 23, 2020Updated 6 years ago
daveredrum / Scan2Cap
View on GitHub
[CVPR 2021] Scan2Cap: Context-aware Dense Captioning in RGB-D Scans
☆106Sep 6, 2022Updated 3 years ago
fkenghagho / RobotVQA
View on GitHub
RobotVQA is a project that develops a Deep Learning-based Cognitive Vision System to support household robots' perception while they perf…
☆18Jul 26, 2024Updated last year
princeton-vl / SimpleView
View on GitHub
Official Code for ICML 2021 paper "Revisiting Point Cloud Shape Classification with a Simple and Effective Baseline"
☆163Dec 12, 2024Updated last year
HaolinLiu97 / Refer-it-in-RGBD
View on GitHub
Repository of our paper 'Refer-it-in-RGBD' in CVPR 2021
☆42May 24, 2024Updated 2 years ago
NUAAXQ / MLCVNet
View on GitHub
[CVPR 2020] MLCVNet: Multi-Level Context VoteNet for 3D Object Detection
☆123Nov 18, 2021Updated 4 years ago
FatemehShiri / Spatial-MM
View on GitHub
☆12Jan 10, 2025Updated last year
jlevin2 / twitch-scraper
View on GitHub
Python script that pulls from twitch API and sends email alerts
☆10Oct 20, 2017Updated 8 years ago
Bare Metal GPUs on DigitalOcean Gradient AI • Ad
Purpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
jiegec / blender-scripts
View on GitHub
Some useful Blender scripts
☆13Jan 15, 2025Updated last year
zyang-ur / SAT
View on GitHub
SAT: 2D Semantics Assisted Training for 3D Visual Grounding, ICCV 2021 (Oral)
☆32Sep 29, 2021Updated 4 years ago
zhangq327 / ARC
View on GitHub
Official Code for ICLR2022 Paper: Chaos is a Ladder: A New Theoretical Understanding of Contrastive Learning via Augmentation Overlap
☆28Sep 28, 2025Updated 9 months ago
amjltc295 / pytorch-golden-template
View on GitHub
PyTorch Golden Template (under development)
☆13Jan 14, 2020Updated 6 years ago
nmheim / NeuralArithmetic.jl
View on GitHub
Collection of layers that can perform arithmetic operations
☆12Aug 17, 2021Updated 4 years ago
LiuAmber / RAHF
View on GitHub
[ACL 2024 main] Aligning Large Language Models with Human Preferences through Representation Engineering (https://aclanthology.org/2024.…
☆28Sep 25, 2024Updated last year
Sparkier / luna
View on GitHub
Luna is inspired by Lucid, a framework for Feature Visualization. However, Luna is built on Tensorflow 2, and thus supports modern models…
☆11Aug 17, 2022Updated 3 years ago
leszekhanusz / diffusion-ui-backend
View on GitHub
Backend for the diffusion-ui frontend
☆25Feb 17, 2024Updated 2 years ago
aistairc / Numeracy-600K
View on GitHub
☆12Jun 19, 2025Updated last year
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
sled-group / COMFORT
View on GitHub
[ICLR 2025 Oral] Official Implementation for "Do Vision-Language Models Represent Space and How? Evaluating Spatial Frame of Reference Un…
☆23Oct 24, 2024Updated last year
cshizhe / vil3dref
View on GitHub
Official implementation of Language Conditioned Spatial Relation Reasoning for 3D Object Grounding (NeurIPS'22).
☆67Dec 2, 2022Updated 3 years ago
connorbrinton / polyai-models
View on GitHub
☆14Sep 28, 2020Updated 5 years ago
McGill-NLP / diffusion-itm
View on GitHub
Code and data setup for the paper "Are Diffusion Models Vision-and-language Reasoners?"
☆33Mar 15, 2024Updated 2 years ago
rirolab / LINGO-Space
View on GitHub
[AAAI 2024] An official implementation of the paper "LINGO-Space: Language-Conditioned Incremental Grounding for Space"
☆12Jul 1, 2024Updated 2 years ago
nihalsid / single-view-3d-reconstruction
View on GitHub
☆25Feb 26, 2021Updated 5 years ago
tamangmilan / llama3
View on GitHub
Building Llama 3 from scratch using PyTorch
☆13Sep 1, 2024Updated last year
yinyunie / Pose2Room
View on GitHub
Implementation of ECCV'2022: Pose2Room: Understanding 3D Scenes from Human Activities
☆89Dec 4, 2023Updated 2 years ago
idansc / mrr-ndcg
View on GitHub
☆18Jun 10, 2024Updated 2 years ago
AI Agents on DigitalOcean Gradient AI Platform • Ad
Build production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
lluma / OCID-Ref
View on GitHub
☆12Apr 25, 2024Updated 2 years ago
zhanghm1995 / Awesome-VQGAN
View on GitHub
Collect papers and codes about VQGAN in various Computer Vision tasks
☆10Dec 20, 2022Updated 3 years ago
CurryYuan / X-Trans2Cap
View on GitHub
[CVPR 2022] X-Trans2Cap: Cross-Modal Knowledge Transfer using Transformer for 3D Dense Captioning
☆36Aug 26, 2022Updated 3 years ago
microsoft / multimodal-aligned-recipe-corpus
View on GitHub
☆18Jun 5, 2024Updated 2 years ago
clp-research / slurk
View on GitHub
Slurk (think “slack for mechanical turk”…) is a lightweight and easily extensible chat server built especially for conducting multimodal …
☆15Dec 8, 2023Updated 2 years ago
mbforbes / physical-commonsense
View on GitHub
Do Neural Language Representations Learn Physical Commonsense?
☆22Dec 28, 2021Updated 4 years ago
coldmanck / RVL-BERT
View on GitHub
The official code for "Visual Relationship Detection with Visual-Linguistic Knowledge from Multimodal Representations" (IEEE Access, 2021…
☆18Oct 21, 2022Updated 3 years ago