GesVLA: Gesture-Aware Vision-Language-Action Model with Embedded Representations
☆29May 22, 2026Updated 4 months ago
Alternatives and similar repositories for GesVLA
Users that are interested in GesVLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2026 Poster] Code and Benchmark for "Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vi…☆21Jun 3, 2026Updated 4 months ago
- The code of the paper "Latent Fingerprint Matching via Dense Minutia Descriptor"☆23Aug 27, 2025Updated last year
- [CVPR 2026] AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation☆87May 24, 2026Updated 4 months ago
- ☆22Feb 22, 2024Updated 2 years ago
- Codes, datasets, and synthetic dataset generator about the paper "LiCamPose: Combining Multi-View LiDAR and RGB Cameras for Robust Single…☆17Feb 28, 2026Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆30Aug 18, 2025Updated last year
- A fingerprint recognition framework featuring fixed-length dense descriptors, robust enhancement, and pose-aware alignment. Code for desc…☆22Apr 24, 2026Updated 5 months ago
- [ICCV 2025] IGL-Nav: Incremental 3D Gaussian Localization for Image-goal Navigation☆68Aug 4, 2025Updated last year
- [ICCV 2025] D^3QE: Learning Discrete Distribution Discrepancy-aware Quantization Error for Autoregressive-Generated Image Detection☆17Jul 11, 2026Updated 2 months ago
- ☆54Apr 3, 2025Updated last year
- Implementation of "HumanReg: Self-supervised Non-rigid Registration of Sparse Human Point Cloud" (3DV 2024)☆15Oct 26, 2024Updated last year
- [CVPR 2024] LiDAR-based Person Re-identification☆60Sep 19, 2024Updated 2 years ago
- [CVPR 2026] Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment☆28May 11, 2026Updated 4 months ago
- [T-IFS 2025] Joint Identity Verification and Pose Alignment for Partial Fingerprints☆31Apr 16, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2025, All Strong Accept] TSP3D: Text-guided Sparse Voxel Pruning for Efficient 3D Visual Grounding☆255Jul 27, 2026Updated 2 months ago
- F2F-AP: Flow-to-Future Asynchronous Policy for Real-time Dynamic Manipulation☆17Sep 5, 2026Updated 3 weeks ago
- Kinetix simulation for Legato: Learning Native Continuation for Action Chunking Flow Policies (RSS 2026)☆35Jun 3, 2026Updated 4 months ago
- Project Page for GaussianFormer☆24May 30, 2024Updated 2 years ago
- [CVPR 2022] Back to Reality: Weakly-supervised 3D Object Detection with Shape-guided Label Enhancement☆44Mar 5, 2024Updated 2 years ago
- [CoRL 2025] GC-VLN: Instruction as Graph Constraints for Training-free Vision-and-Language Navigation☆84Jun 21, 2026Updated 3 months ago
- UniTacHand: Unified Spatio-Tactile Representation for Human-to-Dexterous-Hand Skill Transfer☆26Dec 25, 2025Updated 9 months ago
- Code for Implicit Visual Geometry Transformer (IVGT)☆63May 27, 2026Updated 4 months ago
- ☆23Mar 7, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆22Apr 19, 2026Updated 5 months ago
- OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation☆60Aug 10, 2026Updated last month
- Daily AI/Robotics paper briefings (VLA, World Model, Physical AI)☆24Aug 11, 2026Updated last month
- [ICML 2026] ResVLA: From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges☆29Jun 1, 2026Updated 4 months ago
- ☆132May 10, 2026Updated 4 months ago
- PhyGaP: Physically-Grounded Gaussians with Polarization Cues(CVPR 2026 Oral)☆18May 21, 2026Updated 4 months ago
- ☆52Aug 28, 2026Updated last month
- Official repo for the paper "You Only Gaussian Once: Controllable 3D Gaussian Splatting for Ultra-Densely Sampled Scenes"☆25May 15, 2026Updated 4 months ago
- [NeurIPS 2026] FlowHOI: Flow-based Semantics-Grounded Generation of Hand-Object Interactions for Dexterous Robot Manipulation☆41Sep 25, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation of “Signal Structure-Aware Gaussian Splatting for Large-Scale Scene Reconstruction” (ICLR 2026).☆23Aug 18, 2026Updated last month
- ☆85Jul 24, 2026Updated 2 months ago
- [IEEE TIFS under review] TOPIC: IFViT: Interpretable Fixed-Length Representation for Fingerprint Matching via Vision Transformer☆13Apr 9, 2024Updated 2 years ago
- iComMa: Inverting 3D Gaussian Splatting for Camera Pose Estimation via Comparing and Matching☆135Apr 11, 2024Updated 2 years ago
- ☆19Updated this week
- Curved Projection Reformation☆13Nov 6, 2023Updated 2 years ago
- [CoRL 2026] Official code for Sim-and-Human Co-training for Data-Efficient and Scene-Generalizable Bimanual Manipulation.☆36Sep 21, 2026Updated last week