A list of Offline to Online RL papers (continually updated)
☆103Jul 22, 2026Updated last month
Alternatives and similar repositories for awesome-offline-to-online-RL-papers
Users that are interested in awesome-offline-to-online-RL-papers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [KDD 2023] Causal Inference via Style Transfer for Out-of-distribution Generalisation☆29Feb 29, 2024Updated 2 years ago
- official implementation for our paper Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning (NeurIPS 2023)☆124Jul 31, 2024Updated 2 years ago
- ☆13Apr 16, 2024Updated 2 years ago
- This repository hosts the codebase corresponding to our paper, published at Expert Systems With Applications, titled 'Class-Incremental L…☆14Jun 11, 2024Updated 2 years ago
- Author's PyTorch implementation of ICML'23 paper "Policy Regularization with Dataset Constraint for Offline Reinforcement Learning" for D…☆17Nov 8, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code release for "Supported Policy Optimization for Offline Reinforcement Learning" (NeurIPS 2022), https://arxiv.org/abs/2202.06239☆22Jun 24, 2023Updated 3 years ago
- ☆66Jan 30, 2026Updated 7 months ago
- The official implementation of flow Q-learning (FQL)☆332Jul 21, 2025Updated last year
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆16Feb 10, 2024Updated 2 years ago
- ☆16Apr 14, 2026Updated 4 months ago
- ☆22May 27, 2024Updated 2 years ago
- Advantage weighted Actor Critic for Offline RL☆53Aug 27, 2022Updated 4 years ago
- Decoupled Q-Chunking☆75May 3, 2026Updated 3 months ago
- code for paper "Entropy-regularized Diffusion Policy with Q-Ensembles for Offline Reinforcement Learning"☆21Feb 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [WACV 2024] Domain Generalisation via Risk Distribution Matching☆23Sep 19, 2024Updated last year
- [ICLR 2023 Oral] The official implementation of SQL and EQL in "Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Reg…☆46Jul 27, 2023Updated 3 years ago
- Causal Discovery via Bayesian Optimization (DrBO) - ICLR 2025☆24Apr 13, 2025Updated last year
- [ICML2026] Official JAX code for Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying☆17Jul 3, 2026Updated last month
- [CVPR 2025] h-Edit: Effective and Flexible Diffusion-Based Editing via Doob’s h-Transform☆79Jun 11, 2025Updated last year
- The official implementation of Value Flows☆56Feb 27, 2026Updated 6 months ago
- Author's implementation of ReBRAC, a minimalist improvement upon TD3+BC☆63Aug 3, 2023Updated 3 years ago
- [ICML 2024] The offical implementation of A2PR, a simple way to achieve SOTA in offline reinforcement learning with an adaptive advantage…☆34May 31, 2024Updated 2 years ago
- ☆414Feb 13, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- High-quality single-file implementations of SOTA Offline and Offline-to-Online RL algorithms: AWAC, BC, CQL, DT, EDAC, IQL, SAC-N, TD3+BC…☆1,373Aug 3, 2023Updated 3 years ago
- 🔥🔥🔥 Object State Description & Change Detection☆10Apr 6, 2026Updated 4 months ago
- A benchmark for offline goal-conditioned RL and offline RL☆482Jan 14, 2026Updated 7 months ago
- JAX implementation of WSRL and RL baselines | ICLR 2025☆148Feb 26, 2026Updated 6 months ago
- A collection of offline reinforcement learning algorithms.☆211Nov 26, 2024Updated last year
- An index of algorithms for offline reinforcement learning (offline-rl)☆1,075May 23, 2024Updated 2 years ago
- Implementation of Robust Reinforcement Learning using Offline Data [NeurIPS'22]☆25Nov 9, 2024Updated last year
- Accelerating Research in Plasticity-Motivated Deep Reinforcement Learning.☆46Feb 9, 2026Updated 6 months ago
- A curated list of Diffusion Model in RL resources (continually updated)☆1,634May 30, 2026Updated 3 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆402Feb 5, 2026Updated 6 months ago
- Benchmarked implementations of Offline RL Algorithms.☆77Mar 4, 2025Updated last year
- Code Release for floq: Training Critics via Flow-Matching for Scaling Compute In Value-Based RL☆46Apr 7, 2026Updated 4 months ago
- An elegant PyTorch offline reinforcement learning library for researchers.☆393Aug 9, 2026Updated 3 weeks ago
- Anti exploration in offline reinforcement learning☆11May 17, 2021Updated 5 years ago
- Clean single-file implementation of offline RL algorithms in JAX☆184Jun 5, 2026Updated 2 months ago
- Pessimistic Bootstrapping for Uncertainty-Driven Offline Reinforcement Learning☆29Feb 21, 2022Updated 4 years ago