☆19Apr 25, 2016Updated 10 years ago
Alternatives and similar repositories for trpo
Users that are interested in trpo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A parallel version of Trust Region Policy Optimization☆65Mar 6, 2017Updated 9 years ago
- ☆99Aug 15, 2016Updated 10 years ago
- ☆15Sep 5, 2016Updated 9 years ago
- ☆18Mar 5, 2017Updated 9 years ago
- ☆20Apr 27, 2016Updated 10 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Implementation of TRPO and related algorithms☆655May 20, 2018Updated 8 years ago
- Learning to Reinforcement Learn☆11Nov 22, 2022Updated 3 years ago
- Tensorflow Implementation of Multi-Function Recurrent Unit☆23Jun 13, 2016Updated 10 years ago
- the solustion to https://openai.com/requests-for-research☆12Mar 23, 2017Updated 9 years ago
- Docker build file for mosquitto☆14Apr 3, 2021Updated 5 years ago
- Wikipedia navigation environment for OpenAI Gym☆40Apr 2, 2023Updated 3 years ago
- Model-Free Episodic Control☆14Jan 12, 2017Updated 9 years ago
- A working implementation of the Categorical DQN (Distributional RL).☆95Apr 7, 2018Updated 8 years ago
- ☆38Mar 6, 2017Updated 9 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Understanding Short-Horizon Bias in Stochastic Meta-Optimization☆37Mar 8, 2018Updated 8 years ago
- pybullet_animations☆12Nov 13, 2017Updated 8 years ago
- ☆25Oct 22, 2015Updated 10 years ago
- trust region policy optimization base on gym and tensorflow, can run in distribution mode☆15May 6, 2017Updated 9 years ago
- RWA in pytorch☆14May 7, 2017Updated 9 years ago
- ☆28Apr 15, 2017Updated 9 years ago
- Design good curriculums for deep reinforcement learning☆14May 18, 2016Updated 10 years ago
- Playground for reinforcement learning algorithms implemented in TensorFlow☆16Oct 18, 2016Updated 9 years ago
- reinforcement learning. policy gradient. PCL☆37Apr 25, 2017Updated 9 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆36Aug 2, 2016Updated 10 years ago
- imperative programming in TensorFlow☆18Dec 12, 2016Updated 9 years ago
- Repo for code for the NIPS paper entitled "An Architecture for Deep, Hierarchical Generative Models"☆14Oct 27, 2016Updated 9 years ago
- Code for the paper "Curiosity-driven Exploration in Deep Reinforcement Learning via Bayesian Neural Networks"☆347Nov 22, 2018Updated 7 years ago
- Applying metric learning to kin8nm☆16Nov 10, 2014Updated 11 years ago
- Tensorflow Implementation of the (Dual)-Associative Memory GRUs☆18Jun 14, 2016Updated 10 years ago
- Keras implementation of guide actor-critic for continuous control☆11Mar 12, 2018Updated 8 years ago
- Deep reinforcement learning using an asynchronous advantage actor-critic (A3C) model.☆64Mar 10, 2018Updated 8 years ago
- Learning to Avoid Errors in GANs by Input Space Manipulation (Code for paper)☆23Jul 7, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Add-on package to gym, to record sequences of actions, observations, and rewards☆75Apr 2, 2023Updated 3 years ago
- Code for R:SS 2021 paper RMP2: A Structured Composable Policy Class for Robot Learning.☆45Jun 26, 2021Updated 5 years ago
- Some Reinforcement Learning in Python☆115Apr 17, 2017Updated 9 years ago
- Incorporates external dependencies into HTML file using data: URI scheme☆21Nov 17, 2011Updated 14 years ago
- Add-on for OpenAI Gym that supports automatic downloading of user environments.☆46May 20, 2017Updated 9 years ago
- Towards cross-lingual distributed representations without parallel text trained with adversarial autoencoders☆22Aug 11, 2016Updated 10 years ago
- ☆28Apr 28, 2019Updated 7 years ago