Deep reinforcement learning using an asynchronous advantage actor-critic (A3C) model.
☆64Mar 10, 2018Updated 8 years ago
Alternatives and similar repositories for A3C
Users that are interested in A3C are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Use Asynchronous advantage actor-critic algorithm (A3C) to play Flappy Bird using Keras☆39Aug 26, 2017Updated 9 years ago
- Modified tensorflow implementation of 'Asynchronous Methods for Deep Reinforcement Learning'☆21Dec 15, 2016Updated 9 years ago
- Advantage async actor-critic Algorithms (A3C) and Progressive Neural Network implemented by tensorflow.☆121Oct 12, 2016Updated 9 years ago
- Tensorflow implementation of Asynchronous Advantage Actor Critic (A3C) from "Asynchronous Methods for Deep Reinforcement Learning".☆24Apr 20, 2017Updated 9 years ago
- Asynchronous Methods for Deep Reinforcement Learning☆588Aug 9, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆19Apr 25, 2016Updated 10 years ago
- Simple Example A3C Reinforcement Learning Algorithm in Tensorflow☆13May 23, 2017Updated 9 years ago
- ☆28Oct 9, 2017Updated 8 years ago
- Deep Q-Network (DQN) to play classic Atari Games☆11Sep 18, 2017Updated 8 years ago
- ☆99Aug 15, 2016Updated 10 years ago
- Implementation for ACER in tensorflow and sonnet by deepmind☆11Aug 28, 2017Updated 9 years ago
- ☆15May 31, 2017Updated 9 years ago
- Reinforcement learning algorithm implementations and ML experimentation workspace☆45Jun 8, 2019Updated 7 years ago
- Trust Region Policy Optimization with TensorFlow and OpenAI Gym☆364Jun 2, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Tensorflow + OpenAI Gym implementation of Deep Q-Network (DQN), Double DQN (DDQN), Dueling Network and Deep Deterministic Policy Gradient…☆80Feb 14, 2017Updated 9 years ago
- Accompanying code for "Deep Reinforcement Learning that Matters"☆154Sep 22, 2017Updated 8 years ago
- Using Asynchronous Deep Reinforcement Learning to play Flappy Bird from pixel input.☆30May 22, 2017Updated 9 years ago
- TensorFlow A2C to solve Acrobot, with synchronized parallel environments☆35Apr 21, 2018Updated 8 years ago
- Self-driving car in a simulator controlled by a tiny neural network☆27May 3, 2017Updated 9 years ago
- Combining deep learning and reinforcement learning.☆81Jun 6, 2026Updated 2 months ago
- Asynchronous Advantage Actor Critic☆20Aug 15, 2016Updated 10 years ago
- A3C LSTM Atari with Pytorch plus A3G design☆564Apr 18, 2023Updated 3 years ago
- Implementation of Conditionally Shifted Neurons by Munkhdalai et al. (https://arxiv.org/pdf/1712.09926.pdf)☆28Jul 8, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A parallel version of Trust Region Policy Optimization☆65Mar 6, 2017Updated 9 years ago
- Immortal Flappy Bird - train a Flappy that never dies☆25Sep 4, 2017Updated 8 years ago
- Implementations of deep RL papers and random experimentation☆178Apr 7, 2018Updated 8 years ago
- The potential field method has been studied extensively for autonomous mobile robot path planning in the past decade. The basic concept o…☆10Jul 23, 2022Updated 4 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago
- Proximal Asynchronous SAGA☆13Nov 30, 2017Updated 8 years ago
- PyTorch implementation of "Asynchronous advantage actor-critic"☆19Oct 30, 2025Updated 10 months ago
- Tetris OpenAI environment☆27Mar 13, 2019Updated 7 years ago
- Distributed A3C☆34Dec 22, 2017Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch implementation of Advantage async actor-critic Algorithms (A3C) in PyTorch☆113Apr 3, 2017Updated 9 years ago
- Abstract- As yet, the efficiency optimization of the power electronic converters needs to rely on its circuit model, while an inaccurate …☆14Mar 2, 2023Updated 3 years ago
- Deep Attention Recurrent Q-Network☆115Nov 7, 2015Updated 10 years ago
- Implementation of selected reinforcement learning algorithms in Tensorflow. A3C, DDPG, REINFORCE, DQN, etc.☆153May 28, 2023Updated 3 years ago
- Optimal placement of uav through machine learning methods in 4G newtwork☆10May 11, 2020Updated 6 years ago
- planning trajectories for UAVs☆12Mar 21, 2021Updated 5 years ago
- ☆32Apr 27, 2017Updated 9 years ago