Spatial-Aware VLA Pretraining through Visual-Physical Alignment from Human Videos (CVPR 2026)
☆27Dec 16, 2025Updated 9 months ago
Alternatives and similar repositories for VIPA-VLA
Users that are interested in VIPA-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models (ECCV 2026)☆23Jul 2, 2026Updated 3 months ago
- Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin Representation (ICML 2024)☆12Aug 9, 2024Updated 2 years ago
- AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback (NAACL 2024)☆20Aug 9, 2024Updated 2 years ago
- Being-VL-0.5: Unified Multimodal Understanding via Byte-Pair Visual Encoding (ICCV 2025, Highlight)☆54Dec 22, 2025Updated 9 months ago
- A simple tool to help get information in NKU-EAMIS(NKU Education Affairs Management Information System).☆10Jul 27, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models (ECCV 2026)☆38Jun 30, 2026Updated 3 months ago
- UniTacHand: Unified Spatio-Tactile Representation for Human-to-Dexterous-Hand Skill Transfer☆26Dec 25, 2025Updated 9 months ago
- Kinetix simulation for Legato: Learning Native Continuation for Action Chunking Flow Policies (RSS 2026)☆35Jun 3, 2026Updated 4 months ago
- Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills☆63Jun 19, 2025Updated last year
- Being-H is BeingBeyond's family of human-centric embodied foundation models.☆1,150Sep 27, 2026Updated last week
- Simple tool to help find good price on steam market.☆13Jul 14, 2020Updated 6 years ago
- Awesome paper for multi-modal llm with grounding ability☆21Oct 11, 2025Updated 11 months ago
- [ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos☆509Jun 12, 2026Updated 3 months ago
- Fast Fourier Transform Acceleration Algorithm. (Accelerated by CUDA)☆12Jul 8, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent …☆82Jul 29, 2026Updated 2 months ago
- Pi0-VLA Repository of "MotionTrans: Human VR Data Enable Motion-Level Learning for Robotic Manipulation Policies"☆29Mar 9, 2026Updated 7 months ago
- The repository provides code for EgoMAN model and dataset creation scripts.☆38Dec 31, 2025Updated 9 months ago
- ChatEMG: Synthetic Data Generation to Control a Robotic Hand Orthosis for Stroke☆12Jul 2, 2024Updated 2 years ago
- [ECCV 2026] Official implementation of "RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics"☆97Sep 20, 2026Updated 2 weeks ago
- Official Repository of "Transcrib3D: 3D Referring Expression Resolution through Large Language Models" accepted at IROS 2024☆13Mar 30, 2026Updated 6 months ago
- ☆40Mar 8, 2026Updated 7 months ago
- HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos☆372Apr 16, 2026Updated 5 months ago
- Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos (ICML 2026)☆59May 4, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Being-M0.5: A Real-Time Controllable Vision-Language-Motion Model (ICCV 2025)☆38Sep 4, 2025Updated last year
- ☆10Nov 30, 2022Updated 3 years ago
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆293Jul 7, 2026Updated 3 months ago
- Implementation of "HumanReg: Self-supervised Non-rigid Registration of Sparse Human Point Cloud" (3DV 2024)☆15Oct 26, 2024Updated last year
- From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots (NeurIPS 2025, Spotlight)☆79Sep 28, 2025Updated last year
- SGAP-Net: Semantic-Guided Attentive Prototypes Network for Few-Shot Human-Object Interaction Recognition, AAAI2020.☆14Dec 15, 2020Updated 5 years ago
- A light-weight, Eigen-based C++ library for trajectory optimization for legged robots.☆12Feb 23, 2021Updated 5 years ago
- [ICML 2026] Orienting Latent Actions for Video World Modeling☆124Apr 20, 2026Updated 5 months ago
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation (ICCV 2025)☆26Jan 30, 2026Updated 8 months ago
- AAAI 2026 Oral☆21Dec 23, 2025Updated 9 months ago
- ENPIRE is an agentic framework that allows AI coding agents to autonomously self-improve real-world robot policies through closed-loop ph…☆264Sep 9, 2026Updated last month
- Implementation of the Mesh-VQVAE of "VQ-HPS: Human Pose and Shape Estimation in a Vector-Quantized Latent Space" - ECCV 2024☆19Oct 30, 2024Updated last year
- Official implementation for AAAI-26 paper: "Force-Aware 3D Contact Modeling for Stable Grasp Generation"☆17Sep 27, 2026Updated last week
- Cascade xDAWN EEGNet for ERP detection☆19Apr 27, 2024Updated 2 years ago
- DemoGrasp: Universal Dexterous Grasping from a Single Demonstration (ICLR 2026)☆97Feb 14, 2026Updated 7 months ago