Official repository for "Towards Generalist Embodied AI: A Survey on World Models for VLA Agents". This curated list systematically organizes core resources including research papers, foundation models, evaluation metrics, and benchmarks.
☆51Mar 5, 2026Updated 6 months ago
Alternatives and similar repositories for awesome-world-models-for-vla-agents
Users that are interested in awesome-world-models-for-vla-agents are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A curated list of medical world models for clinical prediction and decision support☆20Mar 17, 2026Updated 5 months ago
- Self-CorrectingVLA:OnlineActionRefinementviaSparseWorldImagination☆27Apr 14, 2026Updated 5 months ago
- LIBERO-X Robustness Litmus for Vision-Language-Action Models☆37Apr 28, 2026Updated 4 months ago
- Keyframe-Chaining VLA, resolving non-Markovian ambiguity via Sparse Semantic History☆21Apr 24, 2026Updated 4 months ago
- A MATALB script for extracting Bag-Of-Visual-Word features.☆11Oct 21, 2016Updated 9 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 🧠 Awesome Memory-VLA: A curated list of Visual-Language-Action models with memory☆132Aug 17, 2026Updated 3 weeks ago
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,483Aug 20, 2026Updated 3 weeks ago
- Active Self-Paced Learning for Cost-Effective and Progressive Face Identification☆17Jun 27, 2018Updated 8 years ago
- Hallucination-Aware Multimodal Benchmark for Gastrointestinal Image Analysis with Large Vision Language Models☆25Oct 12, 2025Updated 11 months ago
- ☆79May 26, 2025Updated last year
- Source Code for Online Collective Matrix Factorization Hashing. Reference: Di Wang, Quan Wang, Yaqiang An, Xinbo Gao, and Yumin Tian. 202…☆11Oct 20, 2020Updated 5 years ago
- Code for ICLR 2022 publication: Who Is the Strongest Enemy? Towards Optimal and Efficient Evasion Attacks in Deep RL. https://openreview…☆10Aug 31, 2024Updated 2 years ago
- XenseRobotics Physical-Ai Intelligence Platform☆19Sep 7, 2026Updated last week
- [CVPR Findings 2026] Large Multimodal Models as General In-Context Classifiers☆26Mar 1, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Comprehensive Empirical Study of Vision-Language Pre-trained Model for Supervised Cross-Modal Retrieval☆43Apr 13, 2022Updated 4 years ago
- Memory-Dependent Manipulation Benchmark based on RoboTwin☆212Sep 7, 2026Updated last week
- 多足机器人MPC控制器+仿真环境☆13Feb 15, 2023Updated 3 years ago
- Hexapod Robot Control☆10May 8, 2023Updated 3 years ago
- Unified Sim2Sim and Sim2Real Deployment Framework for Humanoid Robots - Plug and Play☆27Mar 7, 2026Updated 6 months ago
- ☆20Jul 11, 2023Updated 3 years ago
- The official repository of the paper "X as Supervision: Contending with Depth Ambiguity in Unsupervised Monocular 3D Pose Estimation"☆13Jan 22, 2025Updated last year
- ☆23Aug 9, 2025Updated last year
- [ICRA'25] Official code repository of "QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning"☆23Jun 25, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆44Updated this week
- Sim2real robot manipulation utilizing GS modeling☆14Feb 19, 2025Updated last year
- Learning Visual Feature-Based World Models via Residual Latent Action☆47May 11, 2026Updated 4 months ago
- This project is the official implementation of 'DreamOmni3: Scribble-based Editing and Generation''☆41Aug 8, 2026Updated last month
- This code is for ChaLearn LAP Large-scale Continuous Gesture Recognition Challenge (Round 2) @ICCV 2017☆10Oct 21, 2017Updated 8 years ago
- Deep Joint Semantic-Embedding Hashing(IJCAI2018)☆31Jul 20, 2018Updated 8 years ago
- [CoRL 2026] ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing?☆164Jul 30, 2026Updated last month
- Open source community's implementation of the model from "LANGUAGE MODEL BEATS DIFFUSION — TOKENIZER IS KEY TO VISUAL GENERATION"☆16Nov 11, 2024Updated last year
- ☆27Jun 5, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML2026] Decoupled Multimodal Diffusion Transformer for Bimanual Dexterous Manipulation with a Plugin Tactile Adapter☆36Aug 17, 2026Updated 3 weeks ago
- Track 2: Social Navigation☆27Aug 19, 2025Updated last year
- ☆15Jan 4, 2023Updated 3 years ago
- ☆26Oct 9, 2024Updated last year
- code of [CVPR22] CodedVTR: Codebook-based Sparse Voxel Transformer with Geometric Guidance☆18Jul 10, 2022Updated 4 years ago
- [ECCV 2024] EgoPoseFormer: A Simple Baseline for Stereo Egocentric 3D Human Pose Estimation☆33Mar 6, 2026Updated 6 months ago
- Graph Convolutional Neural Network Hashing☆34Jul 20, 2019Updated 7 years ago