[CVPR 2026] Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment
☆27May 11, 2026Updated 3 months ago
Alternatives and similar repositories for Proxy3D
Users that are interested in Proxy3D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for Implicit Visual Geometry Transformer (IVGT)☆62May 27, 2026Updated 2 months ago
- GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations☆29May 22, 2026Updated 3 months ago
- ☆17May 1, 2026Updated 3 months ago
- ☆32Mar 24, 2026Updated 5 months ago
- F2F-AP: Flow-to-Future Asynchronous Policy for Real-time Dynamic Manipulation☆16Apr 7, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Meta-Memory: Retrieving and Integrating Semantic-Spatial Memories for Robot Spatial Reasoning☆16Nov 26, 2025Updated 8 months ago
- Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers☆22May 28, 2026Updated 2 months ago
- [CVPR 2026 Highlight] Clay-to-Stone: Phase-wise 3D Gaussian Splatting for Monocular Articulated Hand-Object Manipulation Modeling☆21Jun 9, 2026Updated 2 months ago
- ☆18Apr 7, 2026Updated 4 months ago
- IROS☆17Aug 10, 2025Updated last year
- ☆17Jul 6, 2021Updated 5 years ago
- [CVPR 2026] AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation☆79May 24, 2026Updated 3 months ago
- OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding☆17Mar 18, 2026Updated 5 months ago
- [ICCV 2021] Official PyTorch implementation for Deep Relational Metric Learning.☆43Jan 11, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ECCV 2026] Official code of GEM: Generative Supervision Helps Embodied Intelligence☆92May 30, 2026Updated 2 months ago
- [ICCV 2025] D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection☆17Jul 11, 2026Updated last month
- LSRM is a SOTA, feed-forward 3D reconstruction model that generates high-fidelity, relightable 3D digital twins from sparse 2D views.☆53Jun 5, 2026Updated 2 months ago
- Official code for paper "How Much 3D Do Video Foundation Models Encode?"