本项目旨在为致力于进入VLA(Vision-Language-Action)领域的算法工程师提供一份全中文、实战导向的学习/面试手册。 不同于通用的 CV/NLP 面试指南,本项目聚焦于 Robotics 特有的挑战
☆433Jul 25, 2026Updated this week
Alternatives and similar repositories for VLA-Handbook
Users that are interested in VLA-Handbook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing☆3,281Updated this week
- 仅需Python基础,从0构建自己的具身智能机器人;从0逐步构建VLA/OpenVLA/SmolVLA/Pi0, 深入理解具身智能☆2,856Updated this week
- We release Evo-RL, the opensource real-world offline RL on So-101 and AgileX PiPER for easier reproduction.☆721Jul 15, 2026Updated last week
- [Lumina具身智能社区] 具身智能技术指南 Embodied-AI-Guide☆15,049Mar 12, 2026Updated 4 months ago
- ☆55Mar 24, 2026Updated 4 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning☆1,794Jan 6, 2026Updated 6 months ago
- Xbotics 社区具身智能学习指南:我们把“具身综述→学习路线→仿真学习→开源实物→人物访谈→公司图谱”串起来,帮助新手和实战者快速定位路径、落地项目与参与开源。☆1,132Jun 28, 2026Updated 3 weeks ago
- A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (…☆3,403Jul 7, 2026Updated 2 weeks ago
- A Survey on Reinforcement Learning of Vision-Language-Action Models for Robotic Manipulation☆806May 18, 2026Updated 2 months ago
- Xbotics具身智能社 区全网工作汇总☆265Updated this week
- RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI☆4,252Updated this week
- Code for paper: Rethinking Transparent Object Grasping: Depth Completion with Monocular Depth Estimation and Instance Mask☆16Mar 26, 2026Updated 4 months ago
- ☆12,978Jun 16, 2026Updated last month
- 用koch复现lerobot—遥操作数据采集—act复现—diffusion model复现—Pi模型复现—视觉大模型☆30May 16, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- An all-in-one VLA engineering platform for embodied AI — from data to real-robot deployment.☆563Updated this week
- Dexbotic: Open-Source Vision-Language-Action Toolbox☆1,261Jun 25, 2026Updated last month
- unofficial implementation of physical intelligence pi06☆24Nov 26, 2025Updated 7 months ago
- Open-source implementation of "RL Token: Bootstrapping Online RL with Vision-Language-Action Models"☆76Apr 16, 2026Updated 3 months ago
- ☆18Updated this week
- A Pragmatic VLA Foundation Model☆1,665Jun 11, 2026Updated last month
- HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos☆454Updated this week
- ☆30Jul 17, 2026Updated last week
- AAAI 2026 Oral☆18Dec 23, 2025Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Running VLA at 30Hz frame rate and 480Hz trajectory frequency☆593Feb 10, 2026Updated 5 months ago
- [PKU EPIC Lab] 面向小白的具身智能入门指南☆1,058Dec 3, 2025Updated 7 months ago
- Code of the paper "Unseen from Seen: Rewriting Observation-Instruction Using Foundation Models for Augmenting Vision-Language Navigation"…☆20Nov 11, 2025Updated 8 months ago
- ☆585Sep 28, 2025Updated 9 months ago
- Lumina Robotics Talent Call | Lumina社区具身智能招贤榜 | A list for Embodied AI / Robotics Jobs (PhD, RA, intern, etc☆1,451Feb 25, 2026Updated 5 months ago
- RoboTwin 2.0 Offical Repo☆2,629Updated this week
- ☆118Jun 20, 2026Updated last month
- This is a repository for RobustDexGrasp, which achieves robust dexterous grasping of 500+ unseen objects with random poses from single-vi…☆162Aug 20, 2025Updated 11 months ago
- Unofficial OpenPI extension experiment to build a more complete OpenPI-style VLA engineering stack: pi0.5 semantics, RTC, pi0.6 RECAP/MEM…☆15Jul 8, 2026Updated 2 weeks ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Building General-Purpose Robots Based on Embodied Foundation Model☆1,191Updated this week
- ☆615Jul 8, 2026Updated 2 weeks ago
- [ICLR 2026] The offical Implementation of "Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model"☆693Jun 10, 2026Updated last month
- [RSS 2026] Causal video-action world model for generalist robot control☆1,672Jul 9, 2026Updated 2 weeks ago
- InternVLA-A1: Unifying Understanding, Generation, and Action for Robotic Manipulation☆511Updated this week
- Evo-1: Lightweight Vision-Language-Action Model with Preserved Semantic Alignment☆339Updated this week
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,211Apr 3, 2026Updated 3 months ago