Trust Region Policy Optimization with TensorFlow and OpenAI Gym
☆364Jun 2, 2020Updated 6 years ago
Alternatives and similar repositories for trpo
Users that are interested in trpo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of TRPO and related algorithms☆654May 20, 2018Updated 8 years ago
- ☆99Aug 15, 2016Updated 10 years ago
- rllab is a framework for developing and evaluating reinforcement learning algorithms, fully compatible with OpenAI Gym.☆3,078Jun 10, 2023Updated 3 years ago
- ☆162Jul 21, 2017Updated 9 years ago
- PyTorch implementation of Trust Region Policy Optimization☆449Sep 13, 2018Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Reinforcement Learning with Deep Energy-Based Policies☆438Nov 28, 2023Updated 2 years ago
- A continuous action space version of A3C LSTM in pytorch plus A3G design☆259Oct 11, 2024Updated last year
- reinfore learning tool box, contains trpo, a3c algorithm for continous action space☆41Jan 27, 2018Updated 8 years ago
- TensorFlow implementation of the DDPG algorithm from the paper Continuous Control with Deep Reinforcement Learning (ICLR 2016)☆214Feb 16, 2018Updated 8 years ago
- Continuous control with deep reinforcement learning - Deep Deterministic Policy Gradient (DDPG) algorithm implemented in OpenAI Gym envir…☆276Mar 22, 2018Updated 8 years ago
- Code for the paper "Generative Adversarial Imitation Learning"☆728Nov 22, 2018Updated 7 years ago
- Efficient Batched Reinforcement Learning in TensorFlow☆979Jan 11, 2019Updated 7 years ago
- The Winning Solution for the Learning To Run Challenge 2017☆60Jul 4, 2018Updated 8 years ago
- DEPRECATED: Open-source software for robot simulation, integrated with OpenAI Gym.☆2,169Apr 2, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- OpenAI Baselines: high-quality implementations of reinforcement learning algorithms☆16,759Aug 1, 2024Updated 2 years ago
- Guided Policy Search☆600Feb 9, 2021Updated 5 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago
- [ICML 2017] TensorFlow code for Curiosity-driven Exploration for Deep Reinforcement Learning☆1,484Dec 7, 2022Updated 3 years ago
- A parallel version of Trust Region Policy Optimization☆65Mar 6, 2017Updated 9 years ago
- Implementations of deep RL papers and random experimentation☆178Apr 7, 2018Updated 8 years ago
- Reinforcement learning environments with musculoskeletal models☆946Jan 24, 2022Updated 4 years ago
- Tensorforce: a TensorFlow library for applied reinforcement learning☆3,303Updated this week
- Accompanying code for "Deep Reinforcement Learning that Matters"☆154Sep 22, 2017Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Multitask Environments for RL☆283Aug 23, 2021Updated 5 years ago
- Implementation of 'A Distributional Perspective on Reinforcement Learning' and 'Distributional Reinforcement Learning with Quantile Regre…☆133May 5, 2019Updated 7 years ago
- trust region policy optimization base on gym and tensorflow, can run in distribution mode☆15May 6, 2017Updated 9 years ago
- Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees☆93Sep 13, 2019Updated 6 years ago
- reimplementation of the ddpg algorithm using tensorflow☆37Oct 17, 2016Updated 9 years ago
- Open-source implementations of OpenAI Gym MuJoCo environments for use with the OpenAI Gym Reinforcement Learning Research Platform.☆882Oct 16, 2021Updated 4 years ago
- Soft Actor-Critic☆1,302Nov 29, 2023Updated 2 years ago
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 6 years ago
- ☆119Jul 9, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Hacks for training RL systems from John Schulman's lecture at Deep RL Bootcamp (Aug 2017)☆1,123Oct 13, 2017Updated 8 years ago
- ICML 2018 Self-Imitation Learning☆277Apr 18, 2020Updated 6 years ago
- Tensorflow implementation of Generative Adversarial Imitation Learning(GAIL) with discrete action☆112Nov 14, 2018Updated 7 years ago
- Trust Region Policy Optimization (TRPO) in pure TensorFlow☆18Jun 7, 2018Updated 8 years ago
- Proximal Policy Optimization with TensorFlow and OpenAI Gym☆19Mar 31, 2018Updated 8 years ago
- Replicating "Asynchronous Methods for Deep Reinforcement Learning" (http://arxiv.org/abs/1602.01783)☆408Feb 25, 2017Updated 9 years ago
- Publicly releasable baselines for the Retro contest☆129Nov 22, 2018Updated 7 years ago