[CVPR 2026] HiF-VLA: An efficient, bidirectional spatiotemporal expansion Vision-Language-Action Model
☆78Mar 11, 2026Updated 6 months ago
Alternatives and similar repositories for HiF-VLA
Users that are interested in HiF-VLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AR actor for specialist policy training☆28May 4, 2026Updated 4 months ago
- [ICRA 2026] History-Aware Visuomotor Policy Learning via Point Tracking☆27Jan 10, 2026Updated 8 months ago
- Official implementation of FRAPPE: Infusing World Modeling into Generalist Policies via Multiple Future Representation Alignment☆55Mar 24, 2026Updated 5 months ago
- [ICLR 2026] Code of "MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation"☆346Jun 13, 2026Updated 3 months ago
- Official implementation of ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver.☆278Apr 1, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of Spatial-Forcing: Implicit Spatial Representation Alignment for Vision-language-action Model [ICLR2026]☆287Jul 7, 2026Updated 2 months ago
- 🧠 Awesome Memory-VLA: A curated list of Visual-Language-Action models with memory☆136Aug 17, 2026Updated last month
- ContextVLA: Vision-Language-Action Model with Amortized Multi-Frame Context☆20Nov 5, 2025Updated 10 months ago
- [ICML 2026] ResVLA: From Noise to Intent: Anchoring Generative VLA Policies with Residual Bridges☆29Jun 1, 2026Updated 3 months ago
- ☆34Jun 7, 2026Updated 3 months ago
- 🔥 The first open-sourced diffusion vision-langauge-action model. [ICLR 2026]☆185Mar 12, 2026Updated 6 months ago
- [ICML 2026] Latent Reasoning VLA: Latent Thinking and Prediction for Vision-Language-Action Models☆98May 18, 2026Updated 4 months ago
- [CVPR 2026] Official Implementation for Global Prior Meets Local Consistency: Dual-Memory Augmented Vision-Language-Action Model for Effi…☆34Jul 16, 2026Updated 2 months ago
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICCV 2025] Dense Policy (DSP): Bidirectional Autoregressive Learning of Actions☆79Jan 14, 2026Updated 8 months ago
- [ECCV 2026] Official implementation of CEED-VLA: Consistency Vision-Language-Action Model with Early-Exit Decoding.☆52Sep 15, 2025Updated last year
- [ICLR 2026] General Policy Composition (GPC)☆45May 2, 2026Updated 4 months ago
- AsyncVLA: Asynchronous Flow Matching for Vision-Language-Action Models☆28Nov 29, 2025Updated 9 months ago
- [NeurIPS 2025] DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge☆372Jan 6, 2026Updated 8 months ago
- ☆76Jul 25, 2026Updated last month
- [ICML 2026] This repo is the official implementation of "LangForce : Bayesian Decomposition of Vision Language Action Models via Latent …☆80Jul 29, 2026Updated last month
- LLaVA-VLA: A Simple Yet Powerful Vision-Language-Action Model [ICRA 2026]☆213Mar 12, 2026Updated 6 months ago
- Reinforcing Action Policies by Prophesying☆44Aug 7, 2026Updated last month
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- F1: A Vision Language Action Model Bridging Understanding and Generation to Actions☆199Jan 2, 2026Updated 8 months ago
- Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success☆1,387Sep 9, 2025Updated last year
- A unified robotic manipulation learning framework☆24Sep 4, 2025Updated last year
- WoW (World-Omniscient World Model) is a generative world model trained on 2 million robotic interaction trajectories, designed to imagine…☆168Jan 4, 2026Updated 8 months ago
- [ICLR 2026] Codebase for paper "Geometry-aware 4D Video Generation for Robot Manipulation"☆123Jan 10, 2026Updated 8 months ago
- VLA-RFT: Vision-Language-Action Models with Reinforcement Fine-Tuning☆160Oct 6, 2025Updated 11 months ago
- ☆37Sep 7, 2026Updated 2 weeks ago
- HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model☆356Oct 3, 2025Updated 11 months ago
- [ACM MM'26 Oral] MMaDA-VLA: Large Diffusion Vision-Language-Action Model with Unified Multi-Modal Instruction and Generation☆65May 14, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR2026]AtomicVLA: Unlocking the Potential of Atomic Skill Learning in Robots☆84May 23, 2026Updated 3 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 10 months ago
- Memory-Dependent Manipulation Benchmark based on RoboTwin☆216Updated this week
- [CVPR2026] Official repository of paper "CycleManip: Enabling Cyclic Task Manipulation via Effective Historical Perception and Understand…☆26Feb 21, 2026Updated 7 months ago
- Keyframe-Chaining VLA, resolving non-Markovian ambiguity via Sparse Semantic History☆22Apr 24, 2026Updated 4 months ago
- Score and Distribution Matching Policy: Advanced accelerated Visuomotor Policies via matched distillation☆11May 9, 2025Updated last year
- [CVPR2026] Chain of World: World Model Thinking in Latent Motion☆67Mar 4, 2026Updated 6 months ago