Materials for the Practical Sessions of the Reinforcement Learning Summer School 2019: Bandits, RL & Deep RL (PyTorch).
☆90Aug 21, 2019Updated 6 years ago
Alternatives and similar repositories for rlss-2019
Users that are interested in rlss-2019 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for "Calibrated Model-Based Deep Reinforcement Learning", ICML 2019.☆54May 15, 2019Updated 7 years ago
- Reinforcement learning tutorials using the rlberry library.☆18Jan 9, 2023Updated 3 years ago
- An easy-to-use reinforcement learning library for research and education.☆175May 25, 2026Updated 2 months ago
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- ☆28Dec 29, 2025Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official repository for our paper on "Action Inference by Maximising Evidence: Zero-Shot Imitation from Observation with World Models"☆13Dec 4, 2023Updated 2 years ago
- Learning Action-Value Gradients in Model-based Policy Optimization☆32Sep 7, 2021Updated 4 years ago
- ☆27May 17, 2019Updated 7 years ago
- ☆14Jun 7, 2023Updated 3 years ago
- ☆14May 30, 2019Updated 7 years ago
- Some hard problems for reinforcement learning.☆32Oct 5, 2018Updated 7 years ago
- Code associated with the NeurIPS19 paper "Weighted Linear Bandits in Non-Stationary Environments"☆17Nov 14, 2019Updated 6 years ago
- Multi-agent reinforcement learning environment☆39Jul 9, 2019Updated 7 years ago
- Reward Estimation for Variance Reduction in Deep Reinforcement Learning☆11May 8, 2018Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for SPIBB-DQN and Soft-SPIBB-DQN☆11May 5, 2020Updated 6 years ago
- 算法工程师技术栈学习笔记☆15Aug 22, 2022Updated 3 years ago
- Attempt to create a boilerplate Python package structure up-to-date tools and workflows☆15Dec 17, 2022Updated 3 years ago
- Robustness via Retrying: Closed-Loop Robotic Manipulation with Self-Supervised Learning☆16Nov 7, 2018Updated 7 years ago
- A simple Gridworld environment for Open AI gym☆25Jun 10, 2018Updated 8 years ago
- ☆17Oct 25, 2016Updated 9 years ago
- ☆18Jul 6, 2023Updated 3 years ago
- Count based exploration with the successor representation for Unity ML's Pyramid☆13Jun 19, 2019Updated 7 years ago
- Non-stationary Off-policy Evaluation☆13Nov 8, 2018Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Progress, Notes, Summaries and a lot of Questions on Machine Learning☆55Jan 22, 2020Updated 6 years ago
- Simple gym environments for safety in Reinforcement Learning Research☆18Jul 17, 2024Updated 2 years ago
- Implementation of 'A Convolutional Attention Network for Extreme Summarization of Source Code'☆15Mar 14, 2019Updated 7 years ago
- My notes on reinforcement learning papers☆15Jun 14, 2018Updated 8 years ago
- Reinforcement Learning from Hierarchical Critics☆14Jul 30, 2020Updated 6 years ago
- research and implementations of Deep RL agents and their applications☆58Aug 7, 2026Updated last week
- 🔬 Research Framework for Single and Multi-Players 🎰 Multi-Arms Bandits (MAB) Algorithms, implementing all the state-of-the-art algorith…☆424Jun 19, 2026Updated last month
- Clockwork VAEs in JAX/Flax☆32Jul 16, 2021Updated 5 years ago
- ☆27Oct 25, 2019Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation of Population-Guided Parallel Policy Search for Reinforcement Learning☆22Jan 9, 2020Updated 6 years ago
- Combining Evolutionary Algorithms and deep RL in various ways☆108Nov 17, 2020Updated 5 years ago
- Reinforcement Learning via Latent State Decoding☆29Jun 12, 2023Updated 3 years ago
- An INRIA beamer template☆15Nov 19, 2019Updated 6 years ago
- A pack of control system algorithms implemented in C to be used in embedded systems.☆16Dec 7, 2024Updated last year
- Safe Policy Improvement with Baseline Bootstrapping☆26May 5, 2020Updated 6 years ago
- The code accompaniment for the CoRL 2020 paper: A User's Guide to Calibrating Robotics Simulators (https://arxiv.org/abs/2011.08985), fro…☆30Nov 20, 2020Updated 5 years ago