A curated list of academic papers and resources on Vision-Language-Action (VLA) and World Action Models (WAM)
☆34Sep 12, 2026Updated this week
Alternatives and similar repositories for Awesome-World-Action-Model
Users that are interested in Awesome-World-Action-Model are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation☆71Jul 28, 2026Updated last month
- [ICCV 23] A Simple Vision Transformer for Weakly Semi-supervised 3D Object Detection☆13Apr 12, 2024Updated 2 years ago
- [CVPR 2026] PointTPA: Dynamic Network Parameter Adaptation for 3D Scene Understanding☆35Apr 7, 2026Updated 5 months ago
- [CVPR 2026] When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models☆70Apr 11, 2026Updated 5 months ago
- [ECCV 2024] Make Your ViT-based Multi-view 3D Detectors Faster via Token Compression☆53Sep 21, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2026] Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding☆421Jun 18, 2026Updated 2 months ago
- [NeurIPS 2024 Oral] RG-SAN: Rule-Guided Spatial Awareness Network for End-to-End 3D Referring Expression Segmentation☆20Dec 22, 2024Updated last year
- 钢材表面缺陷检测与分割竞赛的解决方案☆24Nov 12, 2024Updated last year
- Next Forcing: Causal World Modeling with Multi-Chunk Prediction (MCP)☆132Aug 16, 2026Updated 3 weeks ago
- the official code of DriveMonkey☆47Mar 20, 2026Updated 5 months ago
- Awesome GPT-4 with Applications. This is a collection of resources related to GPT-4, including news, official documents, demo and applica…☆20Mar 15, 2023Updated 3 years ago
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,475Aug 20, 2026Updated 3 weeks ago
- Out of Sight but Not Out of Mind: Hybrid Memory for Dynamic Video World Models☆278Jul 23, 2026Updated last month
- ☆49Jun 30, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- URDF-based forward and inverse kinematics helpers for robot arms — pluggable IK backends (Placo, PyRoki, RoboPlan), Viser 3D visualizatio…☆24May 22, 2026Updated 3 months ago
- [ICLR26] ThinkOmni: Lifting Textual Reasoning to Omni-modal Scenarios via Guidance Decoding☆86Mar 20, 2026Updated 5 months ago
- Isaac Lab implementation of AMP(Adversarial Motion Prior) with rl_games☆14Aug 5, 2025Updated last year
- Live 3D viewer and editor for MuJoCo MJCF models, inside VS Code☆28Jul 22, 2026Updated last month
- [CVPR 2025] A Unified Image-Dense Annotation Generation Model for Underwater Scenes☆61Apr 9, 2025Updated last year
- Code repository for "Improving Detection of Small Oriented Objects in Aerial Images", WACV 2023 MaCVi - Best Paper Award☆14May 6, 2023Updated 3 years ago
- Vega: Learning to Drive with Natural Language Instructions☆43Mar 27, 2026Updated 5 months ago
- [ICRA 2026] UniFuture: A 4D Driving World Model for Future Generation and Perception☆165Feb 26, 2026Updated 6 months ago
- 2D laser datasets☆15Jan 4, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆18Updated this week
- [ECCV 26] Video Streaming Thinking☆122Jul 28, 2026Updated last month
- 华中科技大学人工智能与自动化学院课程资源☆67May 21, 2023Updated 3 years ago
- Modify by herochiyou☆18May 17, 2024Updated 2 years ago
- My first project: a smart robot based on ROS with 2D lidar sensors and RGB-D camera☆11Jul 7, 2019Updated 7 years ago
- This is a PyTorch implementation of MCLN proposed by our paper "Multi-branch Collaborative Learning Network for 3D Visual Grounding"(ECCV…☆28Oct 10, 2024Updated last year
- Replication of mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs☆27Apr 13, 2026Updated 5 months ago
- [ICML26] Official Repo for WorldCache: Accelerating World Models for Free via Heterogeneous Token Caching☆43Jul 23, 2026Updated last month
- Official implementation of HEAD CoRL 2025☆27Aug 22, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆38Jul 10, 2026Updated 2 months ago
- Simultaneous Localization and Mapping (SLAM) using Lidar, Kinect RGBD measurements☆10Feb 27, 2019Updated 7 years ago
- [MM2024 Oral] 3D-GRES: Generalized 3D Referring Expression Segmentation☆43Dec 15, 2024Updated last year
- [CVPR 2024] Dynamic Adapter Meets Prompt Tuning: Parameter-Efficient Transfer Learning for Point Cloud Analysis☆171Oct 11, 2024Updated last year
- ☆10Jul 25, 2016Updated 10 years ago
- ☆30Jun 29, 2026Updated 2 months ago
- [ICCV 2025] HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation☆263May 12, 2026Updated 4 months ago