☆104Jun 2, 2026Updated 2 months ago
Alternatives and similar repositories for DIAL
Users that are interested in DIAL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆99Jun 2, 2026Updated 2 months ago
- ☆31Apr 11, 2025Updated last year
- ☆20Aug 2, 2026Updated 3 weeks ago
- [ICCV2025 Oral] Latent Motion Token as the Bridging Language for Learning Robot Manipulation from Videos☆180Oct 1, 2025Updated 10 months ago
- ☆101Jun 23, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICML 2026] ResVLA: From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges☆29Jun 1, 2026Updated 2 months ago
- [ECCV 2026] Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?☆98Jul 13, 2025Updated last year
- Official code and data from DexWM ("World Models Can Leverage Human Videos for Dexterous Manipulation").☆96Jun 23, 2026Updated 2 months ago
- [RSS 2026] LDA-1B: Scaling Latent Dynamics Action Model via Universal Embodied Data Ingestion☆310May 26, 2026Updated 3 months ago
- Diffusion Powers Video Tokenizer for Comprehension and Generation (CVPR 2025)☆87Feb 27, 2025Updated last year
- ☆28Jul 2, 2026Updated last month
- Official Release of "Mixture of Horizons in Action Chunking", ICML 2026☆62May 4, 2026Updated 3 months ago
- Replication of mimic-video: Video-Action Models for Generalizable Robot Control Beyond VLAs☆27Apr 13, 2026Updated 4 months ago
- FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation☆113Jul 15, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing☆3,543Updated this week
- Official codebase for Fast-WAM: Do World Action Models Need Test-time Future Imagination?☆1,392Aug 20, 2026Updated last week
- Github repository for World-VLA-Loop.☆33Feb 25, 2026Updated 6 months ago
- [CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model☆76Mar 11, 2026Updated 5 months ago
- official implementation for our paper Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance (CoRL 2024)☆58Apr 28, 2025Updated last year
- InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy☆427Feb 11, 2026Updated 6 months ago
- [RSS 2026] Causal video-action world model for generalist robot control☆1,820Jul 9, 2026Updated last month
- [ECCV 2026] VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model☆552Updated this week
- [ICRA 2026] VITRA: Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos☆483Jun 12, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Codebase for "Do as I Do: Dexterous Manipulation Data from Everyday Human Videos"☆387Aug 2, 2026Updated 3 weeks ago
- DreamAvoid: Critical-Phase Test-Time Dreaming to Avoid Failures in VLA Policies☆20May 20, 2026Updated 3 months ago
- ICLR 2026 Paper: Ctrl-World☆558Apr 8, 2026Updated 4 months ago
- Cosmos Policy☆858Jan 23, 2026Updated 7 months ago
- [ACL2026 Findings] GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning☆84Jun 23, 2025Updated last year
- Galaxea's open-source VLA repository☆757Aug 13, 2026Updated 2 weeks ago
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 5 months ago
- ☆60Jun 4, 2025Updated last year
- RynnVLA-002: A Unified Vision-Language-Action and World Model☆1,122Dec 2, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [IJCV] EgoPlan-Bench: Benchmarking Multimodal Large Language Models for Human-Level Planning☆87Dec 6, 2024Updated last year
- [CVPR 2026] Official Implementation for Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Effi…☆34Jul 16, 2026Updated last month
- Retargeting of whole-body human motion to humanoid robots for dexterous manipulation of articulated objects.☆36Jan 28, 2026Updated 7 months ago
- ☆48Jun 30, 2026Updated last month
- Code to pretrain, fine-tune, and evaluate DreamZero and run sim & real-world evals☆2,602Apr 19, 2026Updated 4 months ago
- [NeurIPS 2024] AlphaTablets: A Generic Plane Representation for 3D Planar Reconstruction from Monocular Videos☆25Dec 6, 2024Updated last year
- RLinf: Reinforcement Learning Infrastructure for Embodied and Agentic AI☆4,674Updated this week