Code for "Policy Decorator: Model-Agnostic Online Refinement for Large Policy Model"
☆119Oct 24, 2025Updated 11 months ago
Alternatives and similar repositories for policy_decorator
Users that are interested in policy_decorator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- From Imitation to Refinement -- Residual RL for Precise Assembly☆263Dec 2, 2025Updated 9 months ago
- ☆91Aug 4, 2025Updated last year
- This is the official implementation of the paper "ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy".☆371Mar 30, 2026Updated 5 months ago
- JAX implementation of WSRL and RL baselines | ICLR 2025☆150Feb 26, 2026Updated 7 months ago
- ☆155Dec 2, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official implementation of DEMO3☆68Jul 29, 2025Updated last year
- ☆409Feb 5, 2026Updated 7 months ago
- ☆40Apr 1, 2024Updated 2 years ago
- Single-file implementation to advance vision-language-action (VLA) models with reinforcement learning.☆458Nov 8, 2025Updated 10 months ago
- Official implementation for DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)☆229Aug 5, 2025Updated last year
- Learning globally stable dynamical systems policies through imitation. A modification of the original work, focussing on waypoint-based i…☆14Oct 12, 2024Updated last year
- ☆68Jul 15, 2025Updated last year
- Code for "DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks"☆22Apr 26, 2024Updated 2 years ago
- Code for Point Policy: Unifying Observations and Actions with Key Points for Robot Manipulation☆92Jul 21, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)☆295Apr 27, 2026Updated 5 months ago
- Unofficial baselines for ManiSkill, including RL and BC algorithms.☆22Jun 6, 2024Updated 2 years ago
- Official Implementation of the paper RiEMann: Near Real-Time SE(3)-Equivariant Robot Manipulation without Point Cloud Segmentation☆41Jan 13, 2026Updated 8 months ago
- Official implementation of Diffusion Policy Policy Optimization, arxiv 2024☆856Feb 4, 2025Updated last year
- ☆259May 12, 2025Updated last year
- Human-in-the-loop Online Rejection Sampling for Robotic Manipulation☆27Nov 3, 2025Updated 10 months ago
- Coarse-to-fine Q-Network☆59Aug 6, 2024Updated 2 years ago
- [ICLR 2026] SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning☆1,866Jan 6, 2026Updated 8 months ago
- [ICML2026] VLAC: A Vision-Language-Action-Critic Model for Robotic Real-World Reinforcement Learning☆332Sep 18, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The official implementation of "Horizon Reduction Makes RL Scalable"☆205Aug 2, 2025Updated last year
- Official PyTorch Implementation of Unified Video Action Model (RSS 2025)☆414Aug 21, 2026Updated last month
- ☆288Aug 25, 2025Updated last year
- ☆414Feb 13, 2023Updated 3 years ago
- ☆16Feb 10, 2025Updated last year
- Reflective Planning: Vision-Language Models for Multi-Stage Long-Horizon Robotic Manipulation☆182Jul 17, 2025Updated last year
- code for the paper Imitation Learning from Observation with Automatic Discount Scheduling☆13Mar 27, 2024Updated 2 years ago
- CoRL25-"AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies"☆51Aug 15, 2025Updated last year
- Q-Estimation and Q-Gating from BC for RL☆55Sep 8, 2026Updated 2 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A Benchmark for Low-Level Manipulation in Home Rearrangement Tasks☆206May 9, 2026Updated 4 months ago
- ☆1,534Oct 27, 2025Updated 11 months ago
- ☆88May 28, 2024Updated 2 years ago
- [RSS 2025] Gripper Keypose and Object Pointflow as Interfaces for Bimanual Robotic Manipulation☆81Jul 22, 2025Updated last year
- The official implementation of flow Q-learning (FQL)☆335Jul 21, 2025Updated last year
- official implementation for our paper Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning (NeurIPS 2023)☆124Jul 31, 2024Updated 2 years ago
- ☆28Aug 20, 2025Updated last year