A curated list of academic papers and resources on Vision-Language-Action (VLA) and World Action Models (WAM)
☆32Aug 23, 2026Updated this week
Alternatives and similar repositories for Awesome-World-Action-Model
Users that are interested in Awesome-World-Action-Model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation☆71Jul 28, 2026Updated 3 weeks ago
- [ICCV 23] A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection☆13Apr 12, 2024Updated 2 years ago
- [CVPR 2026] PointTPA: Dynamic Network Parameter Adaptation for 3D Scene Understanding☆35Apr 7, 2026Updated 4 months ago
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆68Apr 11, 2026Updated 4 months ago
- [ECCV 2024] Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression☆53Sep 21, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ECCV 2026] Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding☆421Jun 18, 2026Updated 2 months ago
- [NeurIPS 2024 Oral] RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation☆20Dec 22, 2024Updated last year
- the official code of DriveMonkey☆46Mar 20, 2026Updated 5 months ago
- Awesome GPT-4 with Applications. This is a collection of resources related to GPT-4, including news, official documents, demo and applica…☆20Mar 15, 2023Updated 3 years ago
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,338Updated this week
- Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models☆275Jul 23, 2026Updated last month
- ☆47Jun 30, 2026Updated last month
- URDF-based forward and inverse kinematics helpers for robot arms — pluggable IK backends (Placo, PyRoki, RoboPlan), Viser 3D visualizatio …☆23May 22, 2026Updated 3 months ago
- [ICLR26] ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding☆86Mar 20, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Isaac Lab implementation of AMP(Adversarial Motion Prior) with rl_games☆14Aug 5, 2025Updated last year
- Live 3D viewer and editor for MuJoCo MJCF models, inside VS Code☆28Jul 22, 2026Updated last month
- [CVPR 2025] A Unified Image-Dense Annotation Generation Model for Underwater Scenes☆61Apr 9, 2025Updated last year
- Official code repository of Shuffle-R1☆26Feb 23, 2026Updated 6 months ago
- Vega: Learning to Drive with Natural Language Instructions☆43Mar 27, 2026Updated 4 months ago
- A light-weight, Eigen-based C++ library for trajectory optimization for legged robots.☆27Aug 11, 2021Updated 5 years ago
- ☆17Updated this week
- [ECCV 26] Video Streaming Thinking☆119Jul 28, 2026Updated 3 weeks ago
- Code for ICLR 2022 publication: Who Is the Strongest Enemy? Towards Optimal and Efficient Evasion Attacks in Deep RL. https://openreview…☆10Aug 31, 2024Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 华中科技大学人工智能与自动化学院课程资源☆67May 21, 2023Updated 3 years ago
- Replication of mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs☆27Apr 13, 2026Updated 4 months ago
- This is a PyTorch implementation of MCLN proposed by our paper "Multi-branch Collaborative Learning Network for 3D Visual Grounding"(ECCV…☆27Oct 10, 2024Updated last year
- ☆35Jul 10, 2026Updated last month
- [ICML26] Official Repo for WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching☆42Jul 23, 2026Updated last month
- [NeurIPS 2024] A Unified Framework for 3D Scene Understanding☆179Jul 7, 2025Updated last year
- Official implementation of HEAD CoRL 2025☆27Aug 22, 2025Updated last year
- [MM2024 Oral] 3D-GRES: Generalized 3D Referring Expression Segmentation☆43Dec 15, 2024Updated last year
- ☆10Jul 25, 2016Updated 10 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2024] Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis☆172Oct 11, 2024Updated last year
- ☆30Jun 29, 2026Updated last month
- [ICCV 2025] HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation☆260May 12, 2026Updated 3 months ago
- ☆62Apr 8, 2026Updated 4 months ago
- 多足机器人MPC控制器+仿真环境☆13Feb 15, 2023Updated 3 years ago
- A growing collection of manipulation tasks built with mjlab.☆66May 7, 2026Updated 3 months ago
- Motion Prediction Work (AAAI2019)☆24Apr 26, 2023Updated 3 years ago