[ICML 2026] Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models
☆86May 18, 2026Updated 2 months ago
Alternatives and similar repositories for LaRA-VLA
Users that are interested in LaRA-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reshaping Action Error Distributions for Reliable Vision-Language-Action Models☆17Feb 5, 2026Updated 5 months ago
- ☆49May 12, 2026Updated 2 months ago
- [ICML 2025] Rethinking Latent Redundancy in Behavior Cloning: An Information Bottleneck Approach for Robot Manipulation☆52Feb 3, 2026Updated 6 months ago
- [CVPR 2026] Official Implementation for Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Effi…☆28Jul 16, 2026Updated 2 weeks ago
- AR actor for specialist policy training☆26May 4, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model☆75Mar 11, 2026Updated 4 months ago
- Official github repository for Sim-and-Human Co-training for Data-Efficient and Generalizable Robotic Manipulation.☆33Mar 13, 2026Updated 4 months ago
- ☆32Jun 7, 2026Updated last month
- Official Repository for RD-VLA☆42Mar 12, 2026Updated 4 months ago
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆272Jul 7, 2026Updated 3 weeks ago
- 🔥 The first open-sourced diffusion vision-langauge-action model. [ICLR 2026]☆185Mar 12, 2026Updated 4 months ago
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 4 months ago
- ☆74Jun 18, 2026Updated last month
- ActionCodec: What Makes for Good Action Tokenizers☆58Mar 1, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- LAP: Language-Action Pre-Training Enables Zero-Shot Cross Embodiment Transfer☆160May 20, 2026Updated 2 months ago
- Official repository for VCoT-Grasp.☆22Nov 18, 2025Updated 8 months ago
- ☆33May 13, 2026Updated 2 months ago
- StarVLA: A Lego-like Codebase for Vision-Language-Action Model Developing☆3,374Updated this week
- This is the official codebase for paper: Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Acti…☆60Jul 11, 2026Updated 3 weeks ago
- [IROS 2025] ReBot: Scaling Robot Learning with Real-to-Sim-to-Real Robotic Video Synthesis☆26May 17, 2025Updated last year
- A practical toolkit for process-level robot evaluation with Process Reward Models (PRMs).☆127Updated this week
- CVPR2025 | TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation☆43Jun 19, 2026Updated last month
- OpenHelix: An Open-source Dual-System VLA Model for Robotic Manipulation☆389Aug 27, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆21Jul 1, 2026Updated last month
- [ECCV 2026] Official implementation of "RoboTracer: Mastering Spatial Trace with Reasoning in Vision-Language Models for Robotics"☆82Jun 18, 2026Updated last month
- Source-free Domain Generalization☆16Sep 24, 2024Updated last year
- [NeurIPS 2025] Code for BEAST Experiments on CALVIN and LIBERO.☆40Jan 8, 2026Updated 6 months ago
- MME-VLA-Suite☆67May 9, 2026Updated 2 months ago
- [ICML 2026] LaST$_0$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model☆87Apr 30, 2026Updated 3 months ago
- [NeurIPS'25] SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning☆40Oct 14, 2025Updated 9 months ago
- ☆33May 16, 2025Updated last year
- InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy☆418Feb 11, 2026Updated 5 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- LITEN: Learning from Inference Time Execution for VLAs☆27Oct 23, 2025Updated 9 months ago
- LaST-R1☆105May 6, 2026Updated 2 months ago
- MemoryWAM: Efficient World Action Modeling with Persistent Memory☆64Jun 19, 2026Updated last month
- [ICRA 2026] History-Aware Visuomotor Policy Learning via Point Tracking☆27Jan 10, 2026Updated 6 months ago
- ☆25Jun 10, 2026Updated last month
- [ICLR 2026] Code of "MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation"☆312Jun 13, 2026Updated last month
- [ACM MM'26] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆63May 14, 2026Updated 2 months ago