Spatial-Aware VLA Pretraining through Visual-Physical Alignment from Human Videos (CVPR 2026)
☆27Dec 16, 2025Updated 9 months ago
Alternatives and similar repositories for VIPA-VLA
Users that are interested in VIPA-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models (ECCV 2026)☆23Jul 2, 2026Updated 2 months ago
- AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback (NAACL 2024)☆20Aug 9, 2024Updated 2 years ago
- Being-VL-0.5: Unified Multimodal Understanding via Byte-Pair Visual Encoding (ICCV 2025, Highlight)☆54Dec 22, 2025Updated 8 months ago
- Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models (ECCV 2026)☆38Jun 30, 2026Updated 2 months ago
- UniTacHand: Unified Spatio-Tactile Representation for Human-to-Dexterous-Hand Skill Transfer☆26Dec 25, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Kinetix simulation for Legato: Learning Native Continuation for Action Chunking Flow Policies (RSS 2026)☆33Jun 3, 2026Updated 3 months ago
- Being-0: A Humanoid Robotic Agent with Vision-Language Models and Modular Skills☆64Jun 19, 2025Updated last year
- Being-H is BeingBeyond's family of human-centric embodied foundation models.☆1,145Aug 25, 2026Updated 3 weeks ago
- Awesome paper for multi-modal llm with grounding ability☆21Oct 11, 2025Updated 11 months ago
- [ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos☆502Jun 12, 2026Updated 3 months ago
- [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent …☆80Jul 29, 2026Updated last month
- Pi0-VLA Repository of "MotionTrans: Human VR Data Enable Motion-Level Learning for Robotic Manipulation Policies"☆28Mar 9, 2026Updated 6 months ago
- The repository provides code for EgoMAN model and dataset creation scripts.☆36Dec 31, 2025Updated 8 months ago
- ChatEMG: Synthetic Data Generation to Control a Robotic Hand Orthosis for Stroke☆12Jul 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ECCV 2026] Official implementation of "RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics"☆84Jun 18, 2026Updated 3 months ago
- Official Repository of "Transcrib3D: 3D Referring Expression Resolution through Large Language Models" accepted at IROS 2024☆13Mar 30, 2026Updated 5 months ago
- ☆40Mar 8, 2026Updated 6 months ago
- HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos☆363Apr 16, 2026Updated 5 months ago
- Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos (ICML 2026)☆59May 4, 2026Updated 4 months ago
- Being-M0.5: A Real-Time Controllable Vision-Language-Motion Model (ICCV 2025)☆38Sep 4, 2025Updated last year
- ☆10Nov 30, 2022Updated 3 years ago
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆286Jul 7, 2026Updated 2 months ago
- Implementation of "HumanReg: Self-supervised Non-rigid Registration of Sparse Human Point Cloud" (3DV 2024)☆15Oct 26, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SGAP-Net: Semantic-Guided Attentive Prototypes Network for Few-Shot Human-Object Interaction Recognition, AAAI2020.☆14Dec 15, 2020Updated 5 years ago
- From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots (NeurIPS 2025, Spotlight)☆78Sep 28, 2025Updated 11 months ago
- ENPIRE is an agentic framework that allows AI coding agents to autonomously self-improve real-world robot policies through closed-loop ph…☆207Sep 9, 2026Updated last week
- A light-weight, Eigen-based C++ library for trajectory optimization for legged robots.☆12Feb 23, 2021Updated 5 years ago
- [ICML 2026] Orienting Latent Actions for Video World Modeling☆121Apr 20, 2026Updated 4 months ago
- Learning Precise Affordances from Egocentric Videos for Robotic Manipulation (ICCV 2025)☆26Jan 30, 2026Updated 7 months ago
- AAAI 2026 Oral☆21Dec 23, 2025Updated 8 months ago
- Official implementation for AAAI-26 paper: "Force-Aware 3D Contact Modeling for Stable Grasp Generation"☆17Mar 13, 2026Updated 6 months ago
- DemoGrasp: Universal Dexterous Grasping from a Single Demonstration (ICLR 2026)☆91Feb 14, 2026Updated 7 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Human-in-the-loop Online Rejection Sampling for Robotic Manipulation☆27Nov 3, 2025Updated 10 months ago
- Code for [MICCAI 2024] MoreStyle: Relax Low-frequency Constraint of Fourier-based Image Reconstruction in Generalizable Medical Image Seg…☆10Oct 25, 2024Updated last year
- Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces☆86Jun 6, 2025Updated last year
- EO: Open-source Unified Embodied Foundation Model Series☆63Jan 15, 2026Updated 8 months ago
- EVOLVE-VLA: Test-Time Training from Environment Feedback for Vision-Language-Action Models☆89Dec 17, 2025Updated 9 months ago
- CoRL25-"AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies"☆50Aug 15, 2025Updated last year
- Getting Started in Imitation Learning☆13Mar 3, 2025Updated last year