ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding
☆20Aug 8, 2025Updated last year
Alternatives and similar repositories for vigil3d
Users that are interested in vigil3d are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repository from the paper "Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind"☆18Mar 18, 2025Updated last year
- [3DV 2026] Official Implementation of the paper "HSM: Hierarchical Scene Motifs for Multi-Scale Indoor Scene Generation"☆39Dec 2, 2025Updated 8 months ago
- ☆26Jun 8, 2026Updated 2 months ago
- [ACMMM 2025] Official implementation of SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero Shot 3D Visual Grounding☆24Nov 25, 2025Updated 8 months ago
- Diorama: Unleashing Zero-shot Single-view 3D Scene Modeling (ICCV 2025 Highlight)☆55Jan 22, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official repository of the NeurIPS 2023 paper "Learning Environment-Aware Affordance for 3D Articulated Object Manipulation under Occlusi…☆27Nov 19, 2024Updated last year
- [CVPR 24] MaskClustering: View Consensus based Mask Graph Clustering for Open-Vocabulary 3D Instance Segmentation☆131Apr 25, 2024Updated 2 years ago
- [TMLR 2025] Monocular Dynamic Gaussian Splatting: Fast, Brittle, and Scene Complexity Rules☆28Aug 1, 2025Updated last year
- [CVPR 2023] EDA: Explicit Text-Decoupling and Dense Alignment for 3D Visual Grounding☆135Oct 11, 2023Updated 2 years ago
- ☆11Jun 19, 2024Updated 2 years ago
- PyTorch implementation for Learning with Twin Noisy Labels for Visible-Infrared Person Re-Identification (CVPR 2022).☆33Feb 1, 2024Updated 2 years ago
- [CVPR 2025] GPS as a Control Signal for Image Generation☆25Mar 18, 2025Updated last year
- [ICCV 2023] Distilling Coarse-to-fine Semantic Matching Knowledge for Weakly Supervised 3D Visual Grounding☆14Oct 2, 2024Updated last year
- [ECCV2022] D3Net: A Unified Speaker-Listener Architecture for 3D Dense Captioning and Visual Grounding☆44Aug 27, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Neural network for sEMG and A-mode ultrasound based hand gesture recognition☆13Aug 13, 2021Updated 5 years ago
- Transferring Genshin PVs into a freehand style with Diffusion Model.☆10Jun 5, 2024Updated 2 years ago
- [ICLR 2025] SINGAPO: Single Image Controlled Generation of Articulated Parts in Objects☆96Feb 17, 2026Updated 6 months ago
- An official implementation for APNet: Urban-level Scene Segmentation of Aerial Images and Point Clouds☆10Feb 7, 2024Updated 2 years ago
- 😎 up-to-date & curated list of awesome 3D Visual Grounding papers, methods & resources.☆283Jan 14, 2026Updated 7 months ago
- [ICCV 2023] Multi3DRefer: Grounding Text Description to Multiple 3D Objects☆98Mar 26, 2026Updated 4 months ago
- [PR 2025] MSCMNet: Multi-scale Semantic Correlation Mining for Visible-Infrared Person Re-Identification☆17Nov 8, 2024Updated last year
- Code for the ECCV22 paper "Bottom Up Top Down Detection Transformers for Language Grounding in Images and Point Clouds"☆95Jun 9, 2023Updated 3 years ago
- [CVPR 2025] 3D-LLaVA: Towards Generalist 3D LMMs with Omni Superpoint Transformer☆101May 26, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆22Apr 14, 2026Updated 4 months ago
- Official implementation of "Spherical Mask: Coarse-to-Fine 3D Point Cloud Instance Segmentation with Spherical Representation"☆94Apr 28, 2024Updated 2 years ago
- [CVPR 2024] Visual Programming for Zero-shot Open-Vocabulary 3D Visual Grounding☆64Aug 3, 2024Updated 2 years ago
- [ECCV 2022, Oral] OPD: Single-view 3D Openable Part Detection☆36Jul 2, 2026Updated last month
- ☆16Apr 3, 2025Updated last year
- [NeurIPS'24] This repository is the implementation of "SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models"☆337Dec 14, 2024Updated last year
- [ICML 2026] ReVSI: Rebuilding Visual Spatial Intelligence Evaluation for Accurate Assessment of VLM 3D Reasoning☆86Jul 8, 2026Updated last month
- ☆35Nov 17, 2025Updated 9 months ago
- RONO: Robust Discriminative Learning with Noisy Labels for 2D-3D Cross-Modal Retrieval (CVPR 2023, PyTorch Code)☆23Mar 11, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- This is the official implementation of "Deep Fuzzy Multi-view Learning for Reliable Classification" (ICML 2025)☆19Jul 16, 2026Updated last month
- Official PyTorch implementation of SynergyNeRF: "Synergistic Integration of Coordinate Network and Tensorial Feature for Improving NeRFs …☆12Sep 23, 2024Updated last year
- ☆13Oct 9, 2024Updated last year
- A small library of 3D related utilities used in my research.☆10Mar 5, 2022Updated 4 years ago
- Cross-Modal Retrieval with Partially Mismatched Pairs (IEEE TPAMI 2023, PyTorch Code)☆23Sep 17, 2023Updated 2 years ago
- 🚘 가짜연구소 10기 '3D perception for Autonomous Driving'☆16Apr 26, 2025Updated last year
- This is an implementation of the paper "Are We Done with Object-Centric Learning?"☆14Jun 21, 2026Updated last month