Spatial-Aware VLA Pretraining through Visual-Physical Alignment from Human Videos (CVPR 2026)
☆26Dec 16, 2025Updated 7 months ago
Alternatives and similar repositories for VIPA-VLA
Users that are interested in VIPA-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models (ECCV 2026)☆22Jul 2, 2026Updated 2 weeks ago
- AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback (NAACL 2024)☆19Aug 9, 2024Updated last year
- A simple tool to help get information in NKU-EAMIS(NKU Education Affairs Management Information System).☆10Jul 27, 2020Updated 5 years ago
- Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models (ECCV 2026)☆38Jun 30, 2026Updated 3 weeks ago
- UniTacHand: Unified Spatio-Tactile Representation for Human-to-Dexterous-Hand Skill Transfer☆26Dec 25, 2025Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills☆63Jun 19, 2025Updated last year
- ☆34Updated this week
- Simple tool to help find good price on steam market.☆13Jul 14, 2020Updated 6 years ago
- Awesome paper for multi-modal llm with grounding ability☆21Oct 11, 2025Updated 9 months ago
- My Blog (https://www.zhangwp.com).☆30Jan 11, 2024Updated 2 years ago
- [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent …☆72Jun 16, 2026Updated last month
- Pi0-VLA Repository of "MotionTrans: Human VR Data Enable Motion-Level Learning for Robotic Manipulation Policies"☆28Mar 9, 2026Updated 4 months ago
- The repository provides code for EgoMAN model and dataset creation scripts.☆32Dec 31, 2025Updated 6 months ago
- [ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos☆443Jun 12, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆39Mar 8, 2026Updated 4 months ago
- [ECCV 2026] Official implementation of "RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics"☆80Jun 18, 2026Updated last month
- HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos☆302Apr 16, 2026Updated 3 months ago
- Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos (ICML 2026)☆51May 4, 2026Updated 2 months ago
- ☆10Nov 30, 2022Updated 3 years ago
- Being-M0.5: A Real-Time Controllable Vision-Language-Motion Model (ICCV 2025)☆37Sep 4, 2025Updated 10 months ago
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆262Jul 7, 2026Updated 2 weeks ago
- Implementation of "HumanReg: Self-supervised Non-rigid Registration of Sparse Human Point Cloud" (3DV 2024)☆15Oct 26, 2024Updated last year
- Codes, datasets, and synthetic dataset generator about the paper "LiCamPose: Combining Multi-View LiDAR and RGB Cameras for Robust Single…☆17Feb 28, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SGAP-Net: Semantic-Guided Attentive Prototypes Network for Few-Shot Human-Object Interaction Recognition, AAAI2020.☆14Dec 15, 2020Updated 5 years ago
- From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots (NeurIPS 2025, Spotlight)☆78Sep 28, 2025Updated 9 months ago
- [ICML 2026] Orienting Latent Actions for Video World Modeling☆116Apr 20, 2026Updated 3 months ago
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation (ICCV 2025)☆26Jan 30, 2026Updated 5 months ago
- Human-centric environment representations from egocentric video☆15Feb 5, 2026Updated 5 months ago
- AAAI 2026 Oral☆18Dec 23, 2025Updated 6 months ago
- Official implementation for AAAI-26 paper: "Force-Aware 3D Contact Modeling for Stable Grasp Generation"☆15Mar 13, 2026Updated 4 months ago
- Human-in-the-loop Online Rejection Sampling for Robotic Manipulation☆27Nov 3, 2025Updated 8 months ago
- Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces☆86Jun 6, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- EO: Open-source Unified Embodied Foundation Model Series☆59Jan 15, 2026Updated 6 months ago
- EVOLVE-VLA: Test-Time Training from Environment Feedback for Vision-Language-Action Models☆87Dec 17, 2025Updated 7 months ago
- CoRL25-"AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies"☆50Aug 15, 2025Updated 11 months ago
- EgoHandICL: Egocentric 3D Hand Reconstruction with In-Context Learning (ICLR 2026)☆32Jan 29, 2026Updated 5 months ago
- Getting Started in Imitation Learning☆13Mar 3, 2025Updated last year
- ☆22Jan 8, 2026Updated 6 months ago
- Reasoning in Space via Grounding in the World (ICLR 2025)☆56Nov 3, 2025Updated 8 months ago