Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)
☆41Jan 28, 2019Updated 7 years ago
Alternatives and similar repositories for MountainCar_DQN_RND
Users that are interested in MountainCar_DQN_RND are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Full Chainer implementation of OpenAI's Reinforcement Learning using Random Network Distillation☆31Apr 15, 2019Updated 7 years ago
- Random Network Distillation pytorch☆263Mar 4, 2019Updated 7 years ago
- DQN based RL agent for Mountain Car☆12Sep 7, 2016Updated 9 years ago
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- Random Network Distillation(RND) algo in Pytorch☆50Feb 26, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Random network distillation on Montezuma's Revenge and Super Mario Bros.☆56May 12, 2025Updated last year
- Code for the paper "Exploration by Random Network Distillation"☆930Oct 1, 2020Updated 5 years ago
- My reproduction of various reinforcement learning algorithms (DQN variants, A3C, DPPO, RND with PPO) in Tensorflow.☆37Mar 24, 2023Updated 3 years ago
- ☆15Nov 22, 2019Updated 6 years ago
- Reinforcement Learning papers on exploration methods.☆19Jun 27, 2021Updated 5 years ago
- 2019 Fall - Game theory and Multi-agent RL Termproject☆10Dec 13, 2019Updated 6 years ago
- Contextual Bandits Action Elimination DQN☆21Jun 25, 2018Updated 8 years ago
- Old and new Reinforcement Learning algorithms run on the GridUniverse ecosystem☆23Feb 3, 2019Updated 7 years ago
- Proximal Policy Optimization(PPO) with Keras Implementation☆17Aug 8, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- pytorch, noisy_distributional_double_dueling_PER_RNN_CNN...CartPole-v1 , Acrobot-v1, MountainCar-v0☆14Mar 19, 2018Updated 8 years ago
- Official implementation of the paper "Approximating two value functions instead of one: towards characterizing a new family of Deep Reinf…☆11Jul 14, 2021Updated 5 years ago
- Deep Reinforcement Learning by using Proximal Policy Optimization and Random Network Distillation in Tensorflow 2 and Pytorch with some e…☆57Nov 10, 2025Updated 9 months ago
- ☆12Jan 3, 2022Updated 4 years ago
- Implementations on OpenAI's Gym☆10Nov 21, 2017Updated 8 years ago
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- Deep Reinforcement Learning Algorithms Implementation in PyTorch☆27Feb 11, 2025Updated last year
- An example and description to Reinforcement Learning DQN model and dataformats for trading☆16Mar 30, 2019Updated 7 years ago
- Logarithmic Reinforcement Learning☆28Apr 7, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Fast and exact spin spherical harmonic transforms☆12Apr 1, 2025Updated last year
- Heuristic Reinforcement Learning☆11Aug 9, 2018Updated 8 years ago
- Co-training for Policy Learning☆13Aug 8, 2019Updated 7 years ago
- Exploring bayesian strategies for approximating optimal actions in POMDPs☆14Jun 27, 2019Updated 7 years ago
- MSc Informatics dissertation project - University of Edinburgh: Curiosity in Multi-Agent Reinforcement Learning☆13Aug 16, 2019Updated 6 years ago
- Coverage path planning with Reinforcement Learning☆11Mar 29, 2022Updated 4 years ago
- An implementation of the AdaOPS (Adaptive Online Packing-based Search), which is an online POMDP Solver used to solve problems defined wi…☆16Nov 16, 2025Updated 8 months ago
- ☆14Apr 8, 2021Updated 5 years ago
- SIMPLE: A Gradient Estimator for $k$-subset Sampling☆12Aug 8, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Method For Establishing Database For Global Value Chain For Parts Procurement☆20Oct 5, 2022Updated 3 years ago
- Software for performing value iteration on partially observable Markov decision processes (POMDPs).☆17Feb 2, 2024Updated 2 years ago
- Collection of Deep Reinforcement Learning Algorithms implemented in PyTorch.☆82Oct 25, 2020Updated 5 years ago
- A pytorch tutorial for DRL(Deep Reinforcement Learning)☆225Apr 24, 2023Updated 3 years ago
- Series of deep reinforcement learning algorithms 🤖☆29Jun 19, 2021Updated 5 years ago
- Deep joint mean and quantile regression for spatio-temporal problems☆16Feb 25, 2020Updated 6 years ago
- yet another reinforcement learning package☆12May 24, 2022Updated 4 years ago