Official PyTorch implementation for ICML 2025 paper: UP-VLA.
☆61Jan 20, 2026Updated 6 months ago
Alternatives and similar repositories for UP-VLA
Users that are interested in UP-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 4 months ago
- ☆33May 16, 2025Updated last year
- ☆62Apr 18, 2025Updated last year
- ✨✨【NeurIPS 2025】Official implementation of BridgeVLA☆193Apr 5, 2026Updated 3 months ago
- ☆25Oct 18, 2025Updated 9 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- The offical repo for paper "VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers" (ICCV 2025)☆133Nov 15, 2025Updated 8 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 6 months ago
- Video Prediction Policy: A Generalist Robot Policy with Predictive Visual Representations https://video-prediction-policy.github.io☆408May 17, 2025Updated last year
- ☆33Sep 25, 2024Updated last year
- Official code for EWMBench: Evaluating Scene, Motion, and Semantic Quality in Embodied World Models☆129Jun 13, 2025Updated last year
- InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy☆419Feb 11, 2026Updated 5 months ago
- ☆475Apr 14, 2026Updated 3 months ago
- Official repository for RSS 25 paper: ArticuBot. Project: https://articubot.github.io/☆41Mar 19, 2026Updated 4 months ago
- [ICLR 2026] Unified Vision-Language-Action Model☆315Oct 15, 2025Updated 9 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- RynnVLA-002: A Unified Vision-Language-Action and World Model☆1,103Dec 2, 2025Updated 7 months ago
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆400Jul 23, 2025Updated last year
- Official repo for From Intention to Execution: Probing the Generalization Boundaries of Vision-Language-Action Models☆33Nov 2, 2025Updated 8 months ago
- [CoRL 2025] Robot Learning from Any Images☆34Nov 11, 2025Updated 8 months ago
- Official code for "Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation" (ICLR2026)☆152Mar 3, 2026Updated 4 months ago
- [ICCV 2025] Dense Policy (DSP): Bidirectional Autoregressive Learning of Actions☆79Jan 14, 2026Updated 6 months ago
- KUDA: Keypoints to Unify Dynamics Learning and Visual Prompting for Open-Vocabulary Robotic Manipulation☆22Apr 23, 2025Updated last year
- 4D-VLA: Spatiotemporal Vision-Language-Action Pretraining with Cross-Scene Calibration. Accepted to NeurIPS 2025.☆58Jan 10, 2026Updated 6 months ago
- ☆20Sep 2, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- code for the paper "ADAPT: Vision-Language Navigation with Modality-Aligned Action Prompts" (CVPR 2022)☆10Jul 17, 2022Updated 4 years ago
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆364Jan 6, 2026Updated 6 months ago
- ICCV2025☆171Dec 10, 2025Updated 7 months ago
- Implementation of VLM4VLA☆165Apr 22, 2026Updated 3 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆32Nov 3, 2025Updated 8 months ago
- AC-DiT: Adaptive Coordination Diffusion Transformer for Mobile Manipulation☆48Feb 23, 2026Updated 5 months ago
- [CVPR 2025 highlight] Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision☆48Dec 2, 2025Updated 7 months ago
- ICCV 2025 | TesserAct: Learning 4D Embodied World Models☆403Aug 4, 2025Updated 11 months ago
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆270Jul 7, 2026Updated 3 weeks ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2026] Disentangled Robot Learning via Separate Forward and Inverse Dynamics Pretraining☆31Apr 26, 2026Updated 3 months ago
- Official PyTorch implementation for NeurIPS 2024 paper: Prediction with Action.☆55Jan 4, 2025Updated last year
- Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation☆178Jul 17, 2025Updated last year
- Embodied Chain of Thought: A robotic policy that reason to solve the task.☆411Apr 5, 2025Updated last year
- TIGeR: Tool-Integrated Geometric Reasoning in Vision-Language Models for Robotics☆23Nov 18, 2025Updated 8 months ago
- [RSS 2026] Ordered Action Tokenization☆103Updated this week
- Official implementation of "OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning"☆236May 30, 2025Updated last year