A comprehensive collection of Vision-Language-Action (VLA) models, benchmarks, and datasets for robotic manipulation and embodied AI research, featuring personally tested reproductions, evaluation environments, and large-scale datasets to serve as a practical guide
☆18Nov 5, 2025Updated 9 months ago
Alternatives and similar repositories for Awesome-VLA
Users that are interested in Awesome-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Vision-Language-Action Optimization with Trajectory Ensemble Voting (ICANN2026)☆27Feb 18, 2026Updated 6 months ago
- ☆48Jun 30, 2026Updated 2 months ago
- A list of papers at the intersection of multimodal learning, embodied AI, and robotics.☆23Jun 28, 2026Updated 2 months ago
- A systematic introduction to Vision Language Action (VLA) models for beginners☆68Mar 23, 2026Updated 5 months ago
- ☆13Apr 17, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of Why Only Text: Empowering Vision-and-Language Navigation with Multi-modal Prompts(IJCAI 2024)☆15Oct 16, 2024Updated last year
- [AAAI 2025] Towards Audio-visual Navigation in Noisy Environments: A Large-scale Benchmark Dataset and An Architecture Considering Multip…☆17May 21, 2026Updated 3 months ago
- Official Repository for the ACM MM 2024 paper "Navigating Beyond Instructions: Vision-and-Language Navigation in Obstructed Environments"☆16May 16, 2025Updated last year
- Implementation (R2R part) for the paper "Iterative Vision-and-Language Navigation"☆18Apr 4, 2024Updated 2 years ago
- The Lottery Ticket Hypothesis for Improving Pretrained Robot Diffusion and Flow Policies☆20May 4, 2026Updated 3 months ago
- 多足机器人MPC控制器+仿真环境☆13Feb 15, 2023Updated 3 years ago
- [CVPR24] OOSTraj: Out-of-Sight Trajectory Prediction With Vision-Positioning Denoising☆16Apr 4, 2024Updated 2 years ago
- My implementation of a scene memory transformer module for reinforcement learning☆14Jun 19, 2019Updated 7 years ago
- 识别工厂中托盘和托盘上的孔☆14Sep 11, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Fisheye Camera Calibration and Undistortion with opencv.☆10Aug 12, 2023Updated 3 years ago
- 《Python编程 从入门到实践》原书配套源代码☆21Oct 18, 2021Updated 4 years ago
- ☆15Aug 8, 2025Updated last year
- 海豚cursor☆15Feb 17, 2025Updated last year
- A comprehensive list of excellent research papers, models, datasets, and other resources on Vision-Language-Action (VLA) models in roboti…☆491Mar 23, 2026Updated 5 months ago
- ☆17Jan 19, 2026Updated 7 months ago
- [WACV 2025, Best Student Paper, Oral] GeoDiffuser: Geometry-Based Image Editing with Diffusion Models☆22Mar 22, 2025Updated last year
- ☆17Dec 23, 2024Updated last year
- This is a Vision-Language-Action Survery☆22Feb 2, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- approximate a mesh with a set of spheres☆35May 17, 2021Updated 5 years ago
- ☆16Feb 22, 2024Updated 2 years ago
- [IROS 2026] Implementation of FailSafe Pipeline in Maniskill Simulator☆24Aug 4, 2026Updated 3 weeks ago
- SLiM: One-shot Quantized Sparse Plus Low-rank Approximation of LLMs (ICML 2025)☆37Nov 28, 2025Updated 9 months ago
- 🎉 [ICLR 2026] All-Day Multi-Scenes Lifelong Vision-and-Language Navigation with Tucker Adaptation☆39Jun 29, 2026Updated 2 months ago
- ☆26Jun 2, 2026Updated 2 months ago
- ☆47May 24, 2024Updated 2 years ago
- ICLR 2026☆46May 29, 2026Updated 3 months ago
- [CVPR'26] Semantic Audio-Visual Navigation in Continuous Environments☆32Jun 23, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official repo for From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models☆34Nov 2, 2025Updated 9 months ago
- A V2V framework that translates human interaction videos into robot manipulation videos.☆24Dec 12, 2025Updated 8 months ago
- The author's Implementation of "Surface-Only Dynamic Deformables using a Boundary Element Method."☆41Sep 12, 2022Updated 3 years ago
- Bridge between LiDAR (Inertial) Odometry and Interactive SLAM☆13Apr 24, 2025Updated last year
- Simple wgpu based SLAM map viewer.☆10May 5, 2020Updated 6 years ago
- Safe Multi-Agent Robosuite benchmark for safe multi-agent reinforcement learning research.☆25Jun 13, 2024Updated 2 years ago
- ☆27Oct 18, 2025Updated 10 months ago