Official PyTorch implementation for ICML 2025 paper: UP-VLA.
☆61Jan 20, 2026Updated 8 months ago
Alternatives and similar repositories for UP-VLA
Users that are interested in UP-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 6 months ago
- ☆33May 16, 2025Updated last year
- ☆66Apr 18, 2025Updated last year
- ✨✨Official implementation of BridgeVLA and BridgeVLA++☆229Aug 13, 2026Updated last month
- ☆28Oct 18, 2025Updated 11 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆134Nov 15, 2025Updated 10 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆125Jan 10, 2026Updated 8 months ago
- Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io☆412May 17, 2025Updated last year
- ☆33Sep 25, 2024Updated 2 years ago
- Official code for EWMBench: Evaluating Scene, Motion, and Semantic Quality in Embodied World Models☆133Jun 13, 2025Updated last year
- InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy☆432Feb 11, 2026Updated 7 months ago
- ☆483Apr 14, 2026Updated 5 months ago
- Official repository for RSS 25 paper: ArticuBot. Project: https://articubot.github.io/☆42Mar 19, 2026Updated 6 months ago
- [ICLR 2026] Unified Vision-Language-Action Model☆322Oct 15, 2025Updated 11 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- RynnVLA-002: A Unified Vision-Language-Action and World Model☆1,136Dec 2, 2025Updated 9 months ago
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆414Aug 21, 2026Updated last month
- Official repo for From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models☆34Nov 2, 2025Updated 10 months ago
- [CoRL 2025] Robot Learning from Any Images☆35Nov 11, 2025Updated 10 months ago
- Official code for "Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation" (ICLR2026)☆156Mar 3, 2026Updated 6 months ago
- [ICCV 2025] Dense Policy (DSP): Bidirectional Autoregressive Learning of Actions☆79Jan 14, 2026Updated 8 months ago
- KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation☆23Apr 23, 2025Updated last year
- 4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.☆59Jan 10, 2026Updated 8 months ago
- ☆23Sep 2, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- code for the paper "ADAPT: Vision-Language Navigation with Modality-Aligned Action Prompts" (CVPR 2022)☆10Jul 17, 2022Updated 4 years ago
- ICCV2025☆172Dec 10, 2025Updated 9 months ago
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆372Jan 6, 2026Updated 8 months ago
- Implementation of VLM4VLA☆172Apr 22, 2026Updated 5 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆34Nov 3, 2025Updated 10 months ago
- AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation☆52Feb 23, 2026Updated 7 months ago
- [CVPR 2025 highlight] Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision☆50Dec 2, 2025Updated 9 months ago
- ICCV 2025 | TesserAct: Learning 4D Embodied World Models☆409Aug 4, 2025Updated last year
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆290Jul 7, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆34Apr 26, 2026Updated 5 months ago
- Official PyTorch implementation for NeurIPS 2024 paper: Prediction with Action.☆56Jan 4, 2025Updated last year
- Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation☆182Jul 17, 2025Updated last year
- Embodied Chain of Thought: A robotic policy that reason to solve the task.☆418Apr 5, 2025Updated last year
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 10 months ago
- [RSS 2026 Finalist] Ordered Action Tokenization☆119Jul 27, 2026Updated 2 months ago
- Official implementation of "OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning"☆239May 30, 2025Updated last year