PyTorch implementation of the Munchausen Reinforcement Learning Algorithms M-DQN and M-IQN
☆46Oct 4, 2020Updated 5 years ago
Alternatives and similar repositories for Munchausen-RL
Users that are interested in Munchausen-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of Soft-Actor-Critic and Prioritized Experience Replay (PER) + Emphasizing Recent Experience (ERE) + Munchausen RL…☆296Feb 24, 2021Updated 5 years ago
- ☆12Feb 21, 2024Updated 2 years ago
- ☆18Sep 7, 2023Updated 3 years ago
- PyTorch Implementation of Implicit Quantile Networks (IQN) for Distributional Reinforcement Learning with additional extensions like PER,…☆95Mar 4, 2023Updated 3 years ago
- Exploration by Random Network Distillation☆15Dec 30, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implicit Distributional Actor Critic☆12Dec 8, 2021Updated 4 years ago
- Code for Diagnosing Bottlenecks in Deep Q-learning. Contains implementations of tabular environments plus solvers.☆17May 14, 2019Updated 7 years ago
- ☆14Feb 14, 2020Updated 6 years ago
- Submission code of UEFDRL team to NeurIPS 2019 MineRL challenge (5th place)☆13Nov 13, 2020Updated 5 years ago
- Code for Optimistic Exploration even with a Pessimistic Initialisation☆14Aug 4, 2020Updated 6 years ago
- Distributed & asynchronous DQN implementation using gRPC and PyTorch.☆10Feb 15, 2021Updated 5 years ago
- DQN-Atari-Agents: Modularized & Parallel PyTorch implementation of several DQN Agents, i.a. DDQN, Dueling DQN, Noisy DQN, C51, Rainbow,…☆122Dec 18, 2020Updated 5 years ago
- Code for the paper Alpha Zero in Continuous Action Space (A0C) (https://arxiv.org/pdf/1805.09613.pdf)☆15Jan 19, 2021Updated 5 years ago
- Smart grid pricing by reinforcement learning☆19Dec 19, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆16Nov 18, 2020Updated 5 years ago
- Various reinforcement learning algorithms written in Jax + Flax☆26Jun 24, 2023Updated 3 years ago
- ☆14Jun 26, 2019Updated 7 years ago
- A collection of RL algorithms written in JAX.☆107Jul 5, 2022Updated 4 years ago
- decision-making processes of human drivers☆15Mar 28, 2024Updated 2 years ago
- Code publication to the paper "Normalized Attention Without Probability Cage"☆17Aug 24, 2026Updated 2 weeks ago
- [CoRL 2020] COG: Connecting New Skills to Past Experience with Offline Reinforcement Learning☆35Oct 28, 2020Updated 5 years ago
- soft q learning and soft actor critic☆16Dec 23, 2018Updated 7 years ago
- Study Group of Model-based RL, 高橋研究室のモデルベース強化学習勉強会のスライドのまとめです☆25Jun 10, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- JAX implementations of various deep reinforcement learning algorithms.☆25Feb 2, 2025Updated last year
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- Map-Elites based on Evolution Strategies☆34Feb 11, 2022Updated 4 years ago
- RLtime is a reinforcement learning library focused on state-of-the-art q-learning algorithms and features☆143Sep 23, 2019Updated 6 years ago
- Author's PyTorch implementation of Randomized Ensembled Double Q-Learning (REDQ) algorithm.☆188Nov 14, 2024Updated last year
- Giving Up Control: Neurons as Reinforcement Learning Agents☆13May 6, 2024Updated 2 years ago
- Alphazero on GPU thanks to CUDA.jl☆34Aug 30, 2021Updated 5 years ago
- ☆11Oct 19, 2020Updated 5 years ago
- P3O paper code☆30Aug 7, 2019Updated 7 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A Reinforcement Learning / Neural Network library, written in Rust.☆21Mar 7, 2021Updated 5 years ago
- Code that can be used to reproduce the experiments in our paper "Estimating Risk and Uncertainty in Deep Reinforcement Learning"☆31Nov 22, 2022Updated 3 years ago
- Jaxplorer is a Jax reinforcement learning (RL) framework for exploring new ideas.☆12Jul 19, 2024Updated 2 years ago
- AR-DAE: Towards Unbiased Neural Entropy Gradient Estimation☆15Jun 22, 2020Updated 6 years ago
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- A beginner's tutorial of reinforcement learning in both Chinese and English. 一份面向初学者的强化学习教程(中英双语)☆13Aug 17, 2023Updated 3 years ago
- Code associated with our paper "Estimating Risk and Uncertainty in Reinforcement Learning"☆11Oct 3, 2023Updated 2 years ago