Reinforcement Learning using Policy Gradient to solve OpenAI Gym games
☆112Dec 13, 2017Updated 8 years ago
Alternatives and similar repositories for openai-gym-policy-gradient
Users that are interested in openai-gym-policy-gradient are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tutorials for the GA Tech OMSCS RLDM class.☆22May 20, 2018Updated 8 years ago
- Re-write of code from Simple Reinforcement Learning with Tensorflow tutorial☆35Jul 28, 2020Updated 5 years ago
- Implementation of DDPG (Modified from the work of Patrick Emami) - Tensorflow (no TFLearn dependency), Ornstein Uhlenbeck noise function,…☆64Apr 27, 2017Updated 9 years ago
- Collection of presentation of my work on various platforms and meetups☆22Feb 2, 2026Updated 5 months ago
- Modular PyTorch implementation of policy gradient methods☆24Nov 15, 2018Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code snippets accompanying the talk "Automatic Differentiation in Haskell."☆19Apr 25, 2020Updated 6 years ago
- Deep RL for portfolio management☆13Aug 31, 2018Updated 7 years ago
- Framework of DataLog Neural Program Synthesis☆27Apr 2, 2019Updated 7 years ago
- ROS-based second-generation command & control system for marine vehicles☆12Apr 15, 2026Updated 3 months ago
- Implementation of VQ-VAE with a GPT-style sampler in the JAX and Haiku ecosystem.☆11Nov 23, 2023Updated 2 years ago
- 레이튼 교수화 최후의 시간여행☆10Apr 19, 2017Updated 9 years ago
- Deep Q-Network (DQN) to play classic Atari Games☆11Sep 18, 2017Updated 8 years ago
- Collision-detection and collision-avoidance navigation demonstration using a feedforward neural network.☆13Nov 4, 2018Updated 7 years ago
- Hands On Reinforcement Learning with Python[Video], Published by Packt☆13Jan 14, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ASTORIA is a framework developed to allow the simulation of attacks and the evaluation of their impact on Smart Grid infrastructures.☆10Feb 5, 2018Updated 8 years ago
- Parallel implementation of DDPG☆13Sep 6, 2017Updated 8 years ago
- The core library of the DFKI multisensor pipeline framework.☆11May 23, 2022Updated 4 years ago
- ☆11Dec 6, 2020Updated 5 years ago
- A mini racetrack world for developing and testing robots with AWS RoboMaker and Gazebo simulations.☆15Sep 8, 2020Updated 5 years ago
- Master Thesis at the Norwegian School of Economics (NHH)☆17Dec 16, 2018Updated 7 years ago
- Minimal implementations of reinforcement learning algorithms by Tensorflow☆29Nov 29, 2017Updated 8 years ago
- Reimplementation of DDPG(Continuous Control with Deep Reinforcement Learning) based on OpenAI Gym + Tensorflow☆574Sep 28, 2021Updated 4 years ago
- OpenAI Gym's LunarLander-v2 Implementation☆42Apr 27, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository for codes of 'Deep Reinforcement Learning'☆218Oct 4, 2019Updated 6 years ago
- ☆16Sep 10, 2017Updated 8 years ago
- The DQN agent which plays breakout-v0 in gym.openai.com☆11Jan 25, 2018Updated 8 years ago
- A tiny implementation of Deep Q Learning, using TensorFlow and OpenAI gym☆93Nov 19, 2021Updated 4 years ago
- A repossitory to discuss about new standard marine messages☆13Oct 15, 2018Updated 7 years ago
- DQN implemented in keras with Dueling Network and Prioritized Experience Replay☆16Nov 21, 2018Updated 7 years ago
- Official code of Nash-DQN for paper: Nash-DQN algorithm for two-player zero-sum Markov games, details see our paper: A Deep Reinforcement…☆22Aug 26, 2022Updated 3 years ago
- ☆13Jan 14, 2020Updated 6 years ago
- ☆28Jun 7, 2019Updated 7 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- FiCo4OMNeT stands for „Fieldbus Communication For OMNeT“. At this point the model consists of two communication technologies – CAN and Fl…☆11Nov 26, 2025Updated 7 months ago
- DQN Pytorch☆16Dec 13, 2021Updated 4 years ago
- Collection of Deep Reinforcement Learning algorithms☆298Mar 19, 2019Updated 7 years ago
- Implementation of Multi-Agent Deep Deterministic Policy Gradients☆39Mar 28, 2018Updated 8 years ago
- An example of clustering applied to finding customer segments.☆20Sep 2, 2017Updated 8 years ago
- Trust Region Policy Optimization with TensorFlow and OpenAI Gym☆364Jun 2, 2020Updated 6 years ago
- Implementation of selected reinforcement learning algorithms in Tensorflow. A3C, DDPG, REINFORCE, DQN, etc.☆153May 28, 2023Updated 3 years ago