Official implementation for DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)
☆219Aug 5, 2025Updated last year
Alternatives and similar repositories for dsrl
Users that are interested in dsrl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation for pi0 steering via DSRL, Steering Your Diffusion Policy with Latent Space Reinforcement Learning (CoRL 2025)☆286Apr 27, 2026Updated 3 months ago
- ☆399Feb 5, 2026Updated 6 months ago
- ☆37Aug 25, 2025Updated 11 months ago
- Official implementation of Diffusion Policy Policy Optimization, arxiv 2024☆845Feb 4, 2025Updated last year
- ☆91Aug 4, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆149Dec 2, 2025Updated 8 months ago
- The official implementation of flow Q-learning (FQL)☆328Jul 21, 2025Updated last year
- [NeurIPS 2025] Flow x RL. "ReinFlow: Fine-tuning Flow Policy with Online Reinforcement Learning". Support VLAs e.g., Pi0, Pi0.5, GR00TN1.…☆354Apr 24, 2026Updated 3 months ago
- Q-learning with Adjoint Matching☆114May 11, 2026Updated 3 months ago
- ☆1,470Oct 27, 2025Updated 9 months ago
- JAX Implementation for Q-Guided Flow and Common RL Baselines☆104Jun 18, 2026Updated last month
- Decoupled Q-Chunking☆74May 3, 2026Updated 3 months ago
- ☆30Jun 30, 2026Updated last month
- Open GVL☆25Dec 1, 2025Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for "Policy Decorator: Model-Agnostic Online Refinement for Large Policy Model"☆118Oct 24, 2025Updated 9 months ago
- [ICLR 2026] General Policy Composition (GPC)☆44May 2, 2026Updated 3 months ago
- [ICML 2025] The Official Implementation of "Efficient Robotic Policy Learning via Latent Space Backward Planning"☆30Dec 15, 2025Updated 7 months ago
- Simplifying diffusion/flow policies by treating action trajectories as flow trajectories☆132Jun 2, 2026Updated 2 months ago
- ☆281Aug 25, 2025Updated 11 months ago
- A benchmark for offline goal-conditioned RL and offline RL