☆99Aug 15, 2016Updated 10 years ago
Alternatives and similar repositories for trpo
Users that are interested in trpo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Apr 25, 2016Updated 10 years ago
- Implementation of TRPO and related algorithms☆654May 20, 2018Updated 8 years ago
- A parallel version of Trust Region Policy Optimization☆65Mar 6, 2017Updated 9 years ago
- Playground for reinforcement learning algorithms implemented in TensorFlow☆16Oct 18, 2016Updated 9 years ago
- ☆20Apr 27, 2016Updated 10 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆15Sep 5, 2016Updated 9 years ago
- trust region policy optimization base on gym and tensorflow, can run in distribution mode☆15May 6, 2017Updated 9 years ago
- ☆162Jul 21, 2017Updated 9 years ago
- Trust Region Policy Optimization with TensorFlow and OpenAI Gym☆364Jun 2, 2020Updated 6 years ago
- Implementations of deep RL papers and random experimentation☆178Apr 7, 2018Updated 8 years ago
- TensorFlow implementation of the DDPG algorithm from the paper Continuous Control with Deep Reinforcement Learning (ICLR 2016)☆214Feb 16, 2018Updated 8 years ago
- some RL algorithms☆19Dec 9, 2016Updated 9 years ago
- PyTorch implementation of Trust Region Policy Optimization☆449Sep 13, 2018Updated 7 years ago
- rllab is a framework for developing and evaluating reinforcement learning algorithms, fully compatible with OpenAI Gym.☆3,077Jun 10, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for the paper "Curiosity-driven Exploration in Deep Reinforcement Learning via Bayesian Neural Networks"☆347Nov 22, 2018Updated 7 years ago
- Code for the paper "Generative Adversarial Imitation Learning"☆728Nov 22, 2018Updated 7 years ago
- Trust Region Policy Optimization with Generalized Advantage Estimator☆16Nov 15, 2018Updated 7 years ago
- "Continuous Deep Q-Learning with Model-based Acceleration" in TensorFlow☆192Jul 20, 2018Updated 8 years ago
- reinforcement learning. policy gradient. PCL☆37Apr 25, 2017Updated 9 years ago
- A working implementation of the Categorical DQN (Distributional RL).☆95Apr 7, 2018Updated 8 years ago
- Training Sonic with RLlib☆61Apr 2, 2023Updated 3 years ago
- pybullet_animations☆12Nov 13, 2017Updated 8 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- TensorFlow implementation of the Value Iteration Networks (NIPS '16) paper☆548Mar 7, 2019Updated 7 years ago
- Guided Policy Search☆600Feb 9, 2021Updated 5 years ago
- Model-Free Episodic Control☆14Jan 12, 2017Updated 9 years ago
- Tensorflow Implementation of Multi-Function Recurrent Unit☆23Jun 13, 2016Updated 10 years ago
- Implementation of algorithms for continuous control (DDPG and NAF).☆311Feb 16, 2021Updated 5 years ago
- Training neural networks with back-prop, feedback-alignment and direct feedback-alignment☆105Jan 15, 2018Updated 8 years ago
- Implement A3C for Mujoco gym envs☆73Nov 2, 2017Updated 8 years ago
- Implementation of the paper [Using Fast Weights to Attend to the Recent Past](https://arxiv.org/abs/1610.06258)☆174Nov 3, 2016Updated 9 years ago
- ☆25Oct 22, 2015Updated 10 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Replicating "Asynchronous Methods for Deep Reinforcement Learning" (http://arxiv.org/abs/1602.01783)☆408Feb 25, 2017Updated 9 years ago
- Reinforcement learning with unsupervised auxiliary tasks☆424Feb 13, 2019Updated 7 years ago
- Asynchronous Advantage Actor Critic☆20Aug 15, 2016Updated 10 years ago
- Asynchronous Methods for Deep Reinforcement Learning☆588Aug 9, 2018Updated 8 years ago
- Deterministic Policy Gradient using torch7☆43Jun 2, 2016Updated 10 years ago
- Implementation of a simple example of Q learning in Torch.☆51Mar 5, 2017Updated 9 years ago
- Robust policy search algorithms which train on model ensembles☆31Oct 26, 2016Updated 9 years ago