Variation of "Asynchronous Methods for Deep Reinforcement Learning" with multiple processes generating experience for agent (Keras + Theano + OpenAI Gym)[1-step Q-learning, n-step Q-learning, A3C]
☆44Feb 27, 2018Updated 8 years ago
Alternatives and similar repositories for async-rl
Users that are interested in async-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Asynchronous Advantage Actor Critic☆20Aug 15, 2016Updated 9 years ago
- Implementation for ACER in tensorflow and sonnet by deepmind☆11Aug 28, 2017Updated 8 years ago
- A3C tensorflow implementation☆11Jul 22, 2018Updated 8 years ago
- Using Pilco algorithm to find a controller for few robotic problems☆43Jul 31, 2015Updated 11 years ago
- Using a paper from Google DeepMind I've developed a new version of the DQN using threads exploration instead of memory replay as explain …☆84Mar 4, 2016Updated 10 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- PyOblige is Python wrapper for OBLIGE - random level generator for Doom☆11Jul 2, 2018Updated 8 years ago
- Model-Free Episodic Control☆14Jan 12, 2017Updated 9 years ago
- Duel_DDQN (Dueling Network Architectures + Double DQN) using Keras☆30Jun 26, 2016Updated 10 years ago
- Asynchronous Methods for Deep Reinforcement Learning☆588Aug 9, 2018Updated 8 years ago
- A3C-LSTM algorithm tested on CartPole OpenAI Gym environment☆48Jul 4, 2018Updated 8 years ago
- A Tensorflow based implementation of "Asynchronous Methods for Deep Reinforcement Learning": https://arxiv.org/abs/1602.01783☆68Oct 28, 2016Updated 9 years ago
- Accompanying repository for Let's make a DQN / A3C series.☆393Sep 4, 2018Updated 7 years ago
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- KEras Reinforcement Learning gYM agents☆290Jul 8, 2017Updated 9 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- Python implementation of tabular asynchronous actor critic☆11May 3, 2016Updated 10 years ago
- OPNET implementations of location-aided routing protocols for MANET☆13Mar 25, 2016Updated 10 years ago
- Deep Reinforcement Learning for Multi Agent Soccer☆16Dec 15, 2016Updated 9 years ago
- an implementation of reinforcement learning problem, stock prices☆10Dec 26, 2016Updated 9 years ago
- Open AI Gym version of Berkeley AI Pacman with images as states☆13May 4, 2018Updated 8 years ago
- Reinforcement learning with unsupervised auxiliary tasks☆424Feb 13, 2019Updated 7 years ago
- AODV in OPNET 14.5☆17Dec 14, 2019Updated 6 years ago
- Estimating stock price correlations using Wikipedia☆25May 11, 2016Updated 10 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- reinforcement learning. policy gradient. PCL☆37Apr 25, 2017Updated 9 years ago
- starter kit for vizdoom2018-singleplayer track☆28Jul 29, 2018Updated 8 years ago
- Multiagent Cooperation and Competition with Deep Reinforcement Learning☆123Nov 26, 2015Updated 10 years ago
- Gym - 32 levels of original Super Mario Bros☆290Dec 21, 2018Updated 7 years ago
- Advantage async actor-critic Algorithms (A3C) and Progressive Neural Network implemented by tensorflow.☆121Oct 12, 2016Updated 9 years ago
- Tensorflow + Keras + OpenAI Gym implementation of 1-step Q Learning from "Asynchronous Methods for Deep Reinforcement Learning"☆1,004Mar 18, 2018Updated 8 years ago
- Deep Learning library for Python. Convnets, recurrent neural networks, and more. Runs on Theano or TensorFlow.☆12Dec 24, 2016Updated 9 years ago
- Application for Math formula detection in image/pdf and then recognition☆13Jan 14, 2025Updated last year
- Use tensorflow2 achieve PPO to play atari game☆13Oct 25, 2019Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SCAN: Learning Abstract Hierarchical Compositional Visual Concepts☆53Oct 29, 2017Updated 8 years ago
- Basic DQN implementation☆226Dec 28, 2017Updated 8 years ago
- An attempt at implementing ideas in "Learning to Transduce with Unbounded Memory" (http://arxiv.org/abs/1506.02516)☆11Jul 27, 2016Updated 10 years ago
- A starter agent that can solve a number of universe environments.☆1,099Apr 7, 2018Updated 8 years ago
- Use Asynchronous advantage actor-critic algorithm (A3C) to play Flappy Bird using Keras☆39Aug 26, 2017Updated 8 years ago
- Implementation of Multi-Agent Deep Deterministic Policy Gradients☆39Mar 28, 2018Updated 8 years ago
- Sharing my solutions to data science hackathons conducted by Analytics Vidhya☆11Apr 29, 2018Updated 8 years ago