Official code for AAAI 2026 paper (One-Step Generative Policies with Q-Learning: A Reformulation of MeanFlow)
☆39Dec 15, 2025Updated 7 months ago
Alternatives and similar repositories for MeanFlowQL
Users that are interested in MeanFlowQL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control☆44Apr 18, 2026Updated 3 months ago
- The official implementation of Value Flows☆55Feb 27, 2026Updated 4 months ago
- [RSS 2026] LPS: Latent Policy Steering through One-Step Flow Policies.☆16Jul 6, 2026Updated 2 weeks ago
- The official implementation of flow Q-learning (FQL)☆321Jul 21, 2025Updated last year
- Decoupled Q-Chunking☆73May 3, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆47Nov 26, 2025Updated 7 months ago
- ☆22Dec 28, 2025Updated 6 months ago
- Guided Flow Policy: Learning from High-Value Actions in Offline RL☆21Apr 28, 2026Updated 2 months ago
- ICRA2026: ABPolicy Asynchronous B-Spline Flow Policy for Real-Time and Smooth Robotic Manipulation☆27Apr 22, 2026Updated 3 months ago
- Code for "Reversal Q-Learning (RQL)" for Flow RL from Prior Data☆32Jun 17, 2026Updated last month
- Q-learning with Adjoint Matching☆109May 11, 2026Updated 2 months ago
- ☆28Sep 10, 2025Updated 10 months ago
- RPent: Agentic Infrastructure for the Physical World☆138Updated this week
- Official Implementation of iMF https://arxiv.org/abs/2512.02012☆331Feb 27, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Sep 4, 2025Updated 10 months ago
- Send command to PX4 using Mavros package☆11Aug 24, 2019Updated 6 years ago
- Projected Overrelaxed Jacobi (JORProx) and Gauss-Seidel (SORProx) GPU implementations.☆14Apr 18, 2026Updated 3 months ago
- official implementation of QVPO☆66Jan 23, 2026Updated 6 months ago
- ☆12Feb 21, 2025Updated last year
- A gym-esque environment for Super Smash Bros. Melee.☆12Aug 13, 2021Updated 4 years ago
- DMPO: Diffusion Model Policy Optimization☆64Jun 9, 2026Updated last month
- ☆10Sep 19, 2021Updated 4 years ago
- Official implementation of SPGrasp: A framework for dynamic grasp synthesis from sparse spatiotemporal prompts.☆20Jun 2, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2026] TwinVLA : Data-Efficient Bimanual Manipulation with Twin Single-Arm Vision-Language-Action Models☆16May 29, 2026Updated last month
- ☆18Jul 14, 2026Updated last week
- Next-gen Foundation Model for Embodied AI☆32Apr 7, 2026Updated 3 months ago
- ☆36Aug 26, 2025Updated 10 months ago
- VLA-GSE: Boosting Parameter Efficient Finetuning in VLA with Generalized and Specialized Experts☆21Jul 11, 2026Updated 2 weeks ago
- Code Release for floq: Training Critics via Flow-Matching for Scaling Compute In Value-Based RL☆46Apr 7, 2026Updated 3 months ago
- Code for "When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?"☆14Dec 19, 2024Updated last year
- [ICLR 2026] [NeurIPS 2025] ViPRA: Video Prediction for Robot Actions☆47Jan 27, 2026Updated 5 months ago
- Official implementation for: Consistency Models as a Rich and Efficient Policy Class for Reinforcement Learning ICLR'24☆27Aug 28, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official repository for the paper "Exploring the Promise and Limits of Real-Time Recurrent Learning" (ICLR 2024)☆13Jun 11, 2025Updated last year
- Simplifying diffusion/flow policies by treating action trajectories as flow trajectories☆128Jun 2, 2026Updated last month
- Self-CorrectingVLA:OnlineActionRefinementviaSparseWorldImagination☆26Apr 14, 2026Updated 3 months ago
- ☆61Apr 8, 2026Updated 3 months ago
- ☆20May 30, 2023Updated 3 years ago
- Official implementation for DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)☆208Aug 5, 2025Updated 11 months ago
- ☆45Jul 1, 2026Updated 3 weeks ago