Model-based Offline Policy Optimization re-implement all by pytorch
☆44Sep 13, 2023Updated 3 years ago
Alternatives and similar repositories for mopo
Users that are interested in mopo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Mar 5, 2024Updated 2 years ago
- Code for MOPO: Model-based Offline Policy Optimization☆191May 17, 2022Updated 4 years ago
- ☆10Mar 11, 2024Updated 2 years ago
- An elegant PyTorch offline reinforcement learning library for researchers.☆393Aug 9, 2026Updated last month
- Implementation of ICLR 2025 paper "Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation"☆18Oct 5, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆35May 24, 2023Updated 3 years ago
- Author's PyTorch implementation of ICML'23 paper "Policy Regularization with Dataset Constraint for Offline Reinforcement Learning" for D…☆17Nov 8, 2024Updated last year
- ☆11Nov 18, 2023Updated 2 years ago
- re-implementation of the offline model-based RL algorithm MOPO in pytorch☆26Feb 28, 2022Updated 4 years ago
- Code for MOBILE: Model-Bellman Inconsistency Penalized Offline Policy Optimization☆22Apr 17, 2024Updated 2 years ago
- Official Codebase for TMLR 2023, Benchmarks and Algorithms for Offline Preference-Based Reward Learning☆20Dec 30, 2022Updated 3 years ago
- Code for demonstration example-task in RUDDER blog☆24May 19, 2020Updated 6 years ago
- Re-implementations of SOTA RL algorithms.☆137Sep 7, 2023Updated 3 years ago
- Implementation of "Reinforcement Learning in Possibly Nonstationary Environments"☆10Mar 10, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Conservative Q learning in Jax☆58Feb 7, 2023Updated 3 years ago
- Implementation of bagging-based ensemble for solar irradiance prediction. Base learners used in ensemble learning is stacked-LSTM☆14Aug 28, 2020Updated 6 years ago
- Official code for "RAMBO: Robust Adversarial Model-Based Offline RL", NeurIPS 2022☆32Jun 2, 2023Updated 3 years ago
- This repository contains the raw data used in "A Multi-Agent Reinforcement Learning Approach to Price and Comfort Optimization in HVAC-Sy…☆11Jan 14, 2022Updated 4 years ago
- rlplot is an easy to use and highly encapsulated RL plot library (including basic error bar lineplot and a wrapper to "rliable").☆33Dec 8, 2023Updated 2 years ago
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- ☆25May 20, 2025Updated last year
- ☆11Mar 15, 2023Updated 3 years ago
- Model-Based Offline Reinforcement Learning☆52Jan 13, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICML 2022] Robust Task Representations for Offline Meta-Reinforcement Learning via Contrastive Learning☆40Aug 17, 2022Updated 4 years ago
- [ICRA'25] H2O+: An Improved Framework for Hybrid Offline-and-Online RL with Dynamics Gaps☆13Apr 10, 2025Updated last year
- Unofficial Pytorch code for "Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models"☆197Dec 8, 2022Updated 3 years ago
- [NeurIPS 2022 Oral] The official implementation of POR in "A Policy-Guided Imitation Approach for Offline Reinforcement Learning"☆58Apr 6, 2023Updated 3 years ago
- Official code for ICLR 2024 paper, SEABO: A Simple Search-Based Method for Offline Imitation Learning☆14Jan 19, 2024Updated 2 years ago
- Flarum WeChat Login extension☆14Oct 19, 2023Updated 2 years ago
- a novel framework based on a physics-informed neural network dubbed as PhysCon that combines the interpretable ability of physical laws a…☆16Jan 4, 2023Updated 3 years ago
- The workers script with deployment for the serverless Halo Download Mirror.☆10Mar 3, 2026Updated 6 months ago
- Format your bibtex (.bib) file to help standardize citations for conference and journal submissions☆14Nov 23, 2025Updated 9 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆18Jan 3, 2020Updated 6 years ago
- ☆17Feb 22, 2021Updated 5 years ago
- A new model-based algorithm for offline inverse reinforcement learning☆15Feb 20, 2023Updated 3 years ago
- ☆13May 21, 2023Updated 3 years ago
- Energy production of photovoltaic (PV) system is heavily influenced by solar irradiance. Accurate prediction of solar irradiance leads to…☆17Aug 30, 2020Updated 6 years ago
- Implementation for POET and POET-X for LLM pretraining☆41Jun 9, 2026Updated 3 months ago
- ☆12May 14, 2024Updated 2 years ago