[CVPR 2026] Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment
☆27May 11, 2026Updated 2 months ago
Alternatives and similar repositories for Proxy3D
Users that are interested in Proxy3D are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for Implicit Visual Geometry Transformer (IVGT)☆62May 27, 2026Updated 2 months ago
- GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations☆29May 22, 2026Updated 2 months ago
- ☆17May 1, 2026Updated 3 months ago
- ☆30Mar 24, 2026Updated 4 months ago
- F2F-AP: Flow-to-Future Asynchronous Policy for Real-time Dynamic Manipulation☆16Apr 7, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Meta-Memory: Retrieving and Integrating Semantic-Spatial Memories for Robot Spatial Reasoning☆16Nov 26, 2025Updated 8 months ago
- Good Token Hunting: A Hitchhiker's Guide to Token Selection for Visual Geometry Transformers☆23May 28, 2026Updated 2 months ago
- [CVPR 2026 Highlight] Clay-to-Stone: Phase-wise 3D Gaussian Splatting for Monocular Articulated Hand-Object Manipulation Modeling☆17Jun 9, 2026Updated last month
- ☆18Apr 7, 2026Updated 3 months ago
- IROS☆17Aug 10, 2025Updated 11 months ago
- ☆17Jul 6, 2021Updated 5 years ago
- [CVPR 2026] AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation☆71May 24, 2026Updated 2 months ago
- OGScene3D: Incremental Open-Vocabulary 3D Gaussian Scene Graph Mapping for Scene Understanding☆16Mar 18, 2026Updated 4 months ago
- [ICCV 2021] Official PyTorch implementation for Deep Relational Metric Learning.☆43Jan 11, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICCV 2025] D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection☆17Jul 11, 2026Updated 3 weeks ago
- LSRM is a SOTA, feed-forward 3D reconstruction model that generates high-fidelity, relightable 3D digital twins from sparse 2D views.☆51Jun 5, 2026Updated last month
- Official code for paper "How Much 3D Do Video Foundation Models Encode?"☆40Mar 24, 2026Updated 4 months ago
- [CVPR 2026 Highlight] SLARM: Streaming and Language-Aligned Reconstruction Model for Dynamic Scenes☆24Jun 9, 2026Updated last month
- 一个面向展示与轻编辑的浏览器端 3DGS 工作台,支持查看、清理、运镜与导出。(A browser-based 3DGS studio for viewing, cleanup editing, shot planning, and video export.)☆22Apr 26, 2026Updated 3 months ago
- Code for paper "ClothTransformer: Unified Latent-Space Transformers for Scalable Cloth Simulation"☆26Jul 17, 2026Updated 2 weeks ago
- PlaneRecTR: Unified Query Learning for 3D Plane Recovery from a Single View☆51Sep 11, 2024Updated last year
- [ECCV 2026] Official implementation of the paper "SegVGGT: Joint 3D Reconstruction and Instance Segmentation from Multi-View Images"☆28May 18, 2026Updated 2 months ago
- [ECCV 2026] PointSplat: Compact Gaussian Splatting via Human-Centric Prediction☆47Jul 14, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- We propose VecSet-Edit a pipeline that enable effecive mesh asset editing using single image.☆38May 4, 2026Updated 3 months ago
- The official repo for “Semantic-guided Semantic Scene Completion”☆18Jul 18, 2024Updated 2 years ago
- Codes for our CVPR 2021 paper "Deep Compositional Metric Learning"☆21Aug 23, 2021Updated 4 years ago
- [CVPR 2026] Official Implementation of "Repurposing 3D Generative Model for Autoregressive Layout Generation"☆60May 19, 2026Updated 2 months ago
- ☆15Jul 13, 2025Updated last year
- [CVPR 2026] 3D-Fixer: Coarse-to-Fine In-place Completion for 3D Scenes from a Single Image☆97Jun 22, 2026Updated last month
- [RSS 2026] FreeOcc: Training-Free Embodied Open-Vocabulary Occupancy Prediction☆118Jul 21, 2026Updated 2 weeks ago
- [CVPR2026] Generalizing Visual Geometry Priors to Sparse Gaussian Occupancy Prediction☆41Jul 16, 2026Updated 2 weeks ago
- ScaleMaster-Dataset☆17May 11, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICML 2026] PhysForge: Generating Physics-Grounded 3D Assets for Interactive Virtual World☆161May 14, 2026Updated 2 months ago
- [SIGGRAPH2026] Official code for SIGGRAPH2026 paper: R-DMesh: Video-Guided 3D Animation via Rectified Dynamic Mesh Flow☆48Jul 18, 2026Updated 2 weeks ago
- TriSplat: Simulation-Ready Feed-Forward 3D Scene Reconstruction☆347Jun 12, 2026Updated last month
- [CVPR 2026] LightSplat: Fast and Memory-Efficient Open-Vocabulary 3D Scene Understanding in Five Seconds☆28Mar 30, 2026Updated 4 months ago
- [AAAI 2026] Official implementation of the paper ”SegDINO3D: 3D Instance Segmentation Empowered by Both Image-Level and Object-Level 2D F…☆69Updated this week
- [ICPR 2026] The official repo for "LayerGS: Decomposition and Inpainting of Layered 3D Human Avatars via 2D Gaussian Splatting"☆15Jan 12, 2026Updated 6 months ago
- TabletopGen: Instance-Level Interactive 3D Tabletop Scene Generation from Text or Single Image☆107Jul 2, 2026Updated last month