[CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model
☆76Mar 11, 2026Updated 5 months ago
Alternatives and similar repositories for HiF-VLA
Users that are interested in HiF-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AR actor for specialist policy training☆26May 4, 2026Updated 3 months ago
- [ICRA 2026] History-Aware Visuomotor Policy Learning via Point Tracking☆27Jan 10, 2026Updated 7 months ago
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 4 months ago
- [ICLR 2026] Code of "MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation"☆317Jun 13, 2026Updated 2 months ago
- Official implementation of ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver.☆271Apr 1, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆281Jul 7, 2026Updated last month
- 🧠 Awesome Memory-VLA: A curated list of Visual-Language-Action models with memory☆111Jul 29, 2026Updated 2 weeks ago
- ContextVLA: Vision-Language-Action Model with Amortized Multi-Frame Context☆20Nov 5, 2025Updated 9 months ago
- [ICML 2026] ResVLA: From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges☆29Jun 1, 2026Updated 2 months ago
- ☆33Jun 7, 2026Updated 2 months ago
- 🔥 The first open-sourced diffusion vision-langauge-action model. [ICLR 2026]☆185Mar 12, 2026Updated 5 months ago
- [ICML 2026] Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models☆88May 18, 2026Updated 2 months ago
- [CVPR 2026] Official Implementation for Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Effi…☆29Jul 16, 2026Updated 3 weeks ago
- [ICLR 2026] General Policy Composition (GPC)☆44May 2, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- [ICCV 2025] Dense Policy (DSP): Bidirectional Autoregressive Learning of Actions☆79Jan 14, 2026Updated 6 months ago
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆52Sep 15, 2025Updated 10 months ago
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆366Jan 6, 2026Updated 7 months ago
- AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models☆28Nov 29, 2025Updated 8 months ago
- [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent …☆75Jul 29, 2026Updated 2 weeks ago
- ☆76Jul 25, 2026Updated 2 weeks ago
- LLaVA-VLA: A Simple Yet Powerful Vision-Language-Action Model [ICRA 2026]☆206Mar 12, 2026Updated 5 months ago
- Reinforcing Action Policies by Prophesying☆41Updated this week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- F1: A Vision Language Action Model Bridging Understanding and Generation to Actions☆199Jan 2, 2026Updated 7 months ago
- Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success☆1,333Sep 9, 2025Updated 11 months ago
- A unified robotic manipulation learning framework☆24Sep 4, 2025Updated 11 months ago
- WoW (World-Omniscient World Model) is a generative world model trained on 2 million robotic interaction trajectories, designed to imagine…☆167Jan 4, 2026Updated 7 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 7 months ago
- Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks☆16Oct 2, 2025Updated 10 months ago
- VLA-RFT: Vision-Language-Action Models with Reinforcement Fine-Tuning☆161Oct 6, 2025Updated 10 months ago
- ☆27Jul 14, 2026Updated 3 weeks ago
- [CVPR2026]AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots☆74May 23, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model☆355Oct 3, 2025Updated 10 months ago
- [ACM MM'26] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆64May 14, 2026Updated 2 months ago
- Memory-Dependent Manipulation Benchmark based on RoboTwin☆194Jul 14, 2026Updated 3 weeks ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 9 months ago
- [CVPR2026] Official repository of paper "CycleManip: Enabling Cyclic Task Manipulation via Effective Historical Perception and Understand…☆25Feb 21, 2026Updated 5 months ago
- Keyframe-Chaining VLA, resolving non-Markovian ambiguity via Sparse Semantic History☆21Apr 24, 2026Updated 3 months ago
- Score and Distribution Matching Policy: Advanced accelerated Visuomotor Policies via matched distillation☆11May 9, 2025Updated last year