Benchmark data for d3rlpy
☆22Nov 28, 2023Updated 2 years ago
Alternatives and similar repositories for d3rlpy-benchmarks
Users that are interested in d3rlpy-benchmarks are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2022] The official implementation of DWBC in "Discriminator-Weighted Offline Imitation Learning from Suboptimal Demonstrations"☆37Jan 5, 2023Updated 3 years ago
- ☆16Oct 5, 2021Updated 4 years ago
- ☆17Nov 1, 2023Updated 2 years ago
- ☆18Apr 11, 2024Updated 2 years ago
- Official implementation of Harnessing Mixed Offline Reinforcement Learning Datasets via Trajectory Reweighting☆16Feb 14, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- An offline deep reinforcement learning library☆1,684Sep 10, 2025Updated 11 months ago
- ☆17May 25, 2023Updated 3 years ago
- The official implementation of "Mind the Gap: Offline Policy Optimization for Imperfect Rewards" (ICLR2023)☆15Mar 3, 2023Updated 3 years ago
- ☆17Dec 30, 2024Updated last year
- Code for NeurIPS 2021 paper "Offline Reinforcement Learning with Reverse Model-based Imagination"☆20Dec 22, 2021Updated 4 years ago
- ☆13Oct 28, 2022Updated 3 years ago
- [REALM25 @ ACL25] - "StateAct" Official Paper Repo (SOTA LLM Agent)☆19Aug 7, 2026Updated 3 weeks ago
- Code for Mildly Conservative Q-learning for Offline Reinforcement Learning (NeurIPS 2022)☆62Apr 29, 2024Updated 2 years ago
- Pessimistic Bootstrapping for Uncertainty-Driven Offline Reinforcement Learning☆29Feb 21, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Re-implementations of SOTA RL algorithms.☆137Sep 7, 2023Updated 2 years ago
- A collection of reference environments for offline reinforcement learning☆1,702Nov 18, 2024Updated last year
- Code for demonstration example-task in RUDDER blog☆24May 19, 2020Updated 6 years ago
- Reinforcement learning with VizDoom platform☆12Apr 18, 2022Updated 4 years ago
- An index of algorithms for offline reinforcement learning (offline-rl)☆1,078May 23, 2024Updated 2 years ago
- ☆15Oct 5, 2020Updated 5 years ago
- Official Implementation of Paper "Learning to Jump: Thinning and Thickening Latent Counts for Generative Modeling" (ICML 2023)☆10Jun 6, 2023Updated 3 years ago
- A systematic design process for a self-organizing neuro-fuzzy Q-network for model-free and offline reinforcement learning.☆11May 29, 2023Updated 3 years ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Codes accompanying the paper "Offline Reinforcement Learning with Value-Based Episodic Memory" (ICLR 2022 https://arxiv.org/abs/2110.0979…☆15Mar 9, 2022Updated 4 years ago
- ☆25Jun 17, 2022Updated 4 years ago
- source code for AAMAS 2023 Imperfect-information Card Game Competition☆13Mar 21, 2024Updated 2 years ago
- Code for "Offline Meta-Reinforcement Learning with Advantage Weighting" [ICML 2021]☆45Nov 30, 2022Updated 3 years ago
- On the model-based stochastic value gradient for continuous reinforcement learning☆58Mar 6, 2026Updated 5 months ago
- Distributed Deep Learning Benchmark Suite☆11Oct 31, 2022Updated 3 years ago
- ☆56Jul 31, 2026Updated last month
- ☆15Oct 2, 2022Updated 3 years ago
- ☆16Jun 1, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for FOCAL Paper Published at ICLR 2021☆55Dec 4, 2023Updated 2 years ago
- Implementation of ``Actor-Critic Alignment for Offline-to-Online Reinforcement Learning''☆13Oct 12, 2023Updated 2 years ago
- Probabilistic Programming in Python. Uses Theano as a backend and includes the NUTS sampler.☆12Apr 19, 2017Updated 9 years ago
- Exploring algorithms in the domain of offline reinforcement learning (REM, Ensemble-DQN, DQN, ...)☆17Jul 7, 2020Updated 6 years ago
- "Hierarchical Reinforcement Learning for Integrated Recommendation" (AAAI 2021) https://ojs.aaai.org/index.php/AAAI/article/view/16580☆58Sep 12, 2021Updated 4 years ago
- Code for ICML2020 "Sequence Generation with Mixed Representations"☆12Jun 27, 2020Updated 6 years ago
- ☆16Sep 16, 2022Updated 3 years ago