A paper list that includes world models or generative video models for embodied agents.
☆27Jan 17, 2025Updated last year
Alternatives and similar repositories for Awesome-embodied-world-model-papers
Users that are interested in Awesome-embodied-world-model-papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code for the NeurIPS'23 paper "3D-Aware Visual Question Answering about Parts, Poses and Occlusions"☆21Oct 17, 2024Updated last year
- [ACL 2023] Code and data for our paper "Measuring Progress in Fine-grained Vision-and-Language Understanding"☆13Jun 11, 2023Updated 3 years ago
- ☆41Feb 3, 2026Updated 7 months ago
- Embodied Agent Interface (EAI): Benchmarking LLMs for Embodied Decision Making (NeurIPS D&B 2024 Oral)☆299Mar 6, 2025Updated last year
- ☆10Jul 5, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Sep 2, 2023Updated 3 years ago
- HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction☆41Sep 15, 2025Updated last year
- Latest Learning-based Embodied Humanoid Robot Methods☆16Sep 22, 2025Updated 11 months ago
- [CVPR 2025 Oral] FluidNexus: 3D Fluid Reconstruction and Prediction from a Single Video☆82Mar 22, 2026Updated 5 months ago
- ☆17May 23, 2023Updated 3 years ago
- Official Code for SGRv2 and SGR.☆33May 20, 2025Updated last year
- [ICML 2025 Oral] Official repo of EmbodiedBench, a comprehensive benchmark designed to evaluate MLLMs as embodied agents.☆343May 30, 2026Updated 3 months ago
- Dreamitate: Real-World Visuomotor Policy Learning via Video Generation (CoRL 2024)☆59Jun 7, 2025Updated last year
- Deep Reinforcement Learning - Implementations and Theory: A path to mastery☆13Nov 21, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- (RA-L 2025) VILP: Imitation Learning with Latent Video Planning☆27Jun 21, 2025Updated last year
- ☆19Jun 26, 2025Updated last year
- Personal Claude Code plugin marketplace☆16Updated this week
- Codebase for VidHal: Benchmarking Hallucinations in Vision LLMs☆14Apr 23, 2026Updated 4 months ago
- Code implementation for: From Virtual Games to Real-World Play☆49Jun 23, 2025Updated last year
- An implementation using NVIDIA Kaolin Wisp and Stable Diffusion that applies the Score Distillation Sampling to the task of Few-shot / si…☆36Jan 31, 2023Updated 3 years ago
- Official repository for "LIV: Language-Image Representations and Rewards for Robotic Control" (ICML 2023)☆136Oct 19, 2023Updated 2 years ago
- 2018研究生推免计算机类高校夏令营时间安排☆12May 14, 2018Updated 8 years ago
- ☆16Mar 8, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆15May 4, 2025Updated last year
- 分别使用KNN和SVM在CIFAR10数据集上进行物体分类任务,后续加入了HOG特征提取对图像做预处理,提高SVM的分类性能。☆15Sep 8, 2021Updated 5 years ago
- ☆42May 3, 2026Updated 4 months ago
- Official Pytorch implementation for LARP: Tokenizing Videos with a Learned Autoregressive Generative Prior (ICLR 2025 Oral).☆108Feb 11, 2025Updated last year
- Official implementation of NeurIPS24 paper "PhyRecon: Physically Plausible Neural Scene Reconstruction"☆173Mar 21, 2025Updated last year
- ☆16Jan 30, 2025Updated last year
- [CVPR 2024] Physical Property Understanding from Language-Embedded Feature Fields☆96Nov 16, 2025Updated 10 months ago
- Benchmark and training code for MindCube: spatial mental modeling in vision-language models from limited views.☆172Aug 23, 2026Updated 3 weeks ago
- ☆14Feb 21, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- THEORY OF SPACE: a benchmark for evaluating whether foundation models can actively explore under partial observability efficiently to bui…☆87Feb 27, 2026Updated 6 months ago
- ☆19Mar 9, 2025Updated last year
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 10 months ago
- [TMLR 2025] The official repository of the paper "Unsupervised Discovery of Object-Centric Neural Fields"☆18Feb 15, 2026Updated 7 months ago
- ☆22Sep 26, 2024Updated last year
- [CVPR'25 Highlight] Official implementation for paper - LeviTor: 3D Trajectory Oriented Image-to-Video Synthesis☆163Apr 15, 2025Updated last year
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆411Aug 21, 2026Updated 3 weeks ago