Proximal Policy Optimization with Stein Control Variates:
☆34Feb 12, 2018Updated 8 years ago
Alternatives and similar repositories for PPO-Stein-Control-Variate
Users that are interested in PPO-Stein-Control-Variate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Stein Variational Policy Gradient for REINFORCE☆18Jul 12, 2017Updated 9 years ago
- ☆162Jul 21, 2017Updated 9 years ago
- Code release for the ICLR paper☆22Jun 13, 2018Updated 8 years ago
- Tutorial on continuous control at Reinforcement Learning Summer School 2017.☆34Jul 3, 2017Updated 9 years ago
- ☆10Apr 2, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This repository contains implementations of the paper, Bayesian Model-Agnostic Meta-Learning.☆20Jan 19, 2023Updated 3 years ago
- PyTorch implementation of Sample Efficient Actor-Critic with Experience Replay(ACER)☆16Oct 7, 2020Updated 5 years ago
- Surprise-based intrinsic motivation for deep reinforcement learning☆21Mar 6, 2017Updated 9 years ago
- AISTATS 2019: Reference-based Adversarial Sampling & Its applications to Soft Q-learning☆15Jan 21, 2019Updated 7 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago
- Model-based reinforcement learning (generative simulator models and planning agents)☆16Mar 13, 2026Updated 5 months ago
- ☆29Nov 21, 2022Updated 3 years ago
- Experiments of amortized stein variational gradient☆17Apr 30, 2017Updated 9 years ago
- TD-VAE in PyTorch☆10May 28, 2019Updated 7 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ICML 2018 Self-Imitation Learning☆277Apr 18, 2020Updated 6 years ago
- ☆86Apr 10, 2021Updated 5 years ago
- Public accompanying repository for Universite de Montreal's IFT 6757: Autnonomous Vehicles, Fall 2019.☆11Jun 21, 2022Updated 4 years ago
- Reinforcement Learning with Deep Energy-Based Policies☆438Nov 28, 2023Updated 2 years ago
- Code for the paper "Evolved Policy Gradients"☆254Nov 22, 2018Updated 7 years ago
- code for the paper "Stein Variational Gradient Descent (SVGD): A General Purpose Bayesian Inference Algorithm"☆423Mar 21, 2024Updated 2 years ago
- A tensorflow implementation of VAE training with Renyi divergence☆32Sep 14, 2016Updated 9 years ago
- ☆14May 15, 2025Updated last year
- PyTorch implementation of Stein Variational Gradient Descent☆49Jun 16, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository contains implementations of the paper, Bayesian Model-Agnostic Meta-Learning.☆59Jul 12, 2019Updated 7 years ago
- Reinforcement learning benchmarking.☆39Oct 22, 2018Updated 7 years ago
- a library for deep reinforcement learning, with applications for navigation☆16Feb 6, 2018Updated 8 years ago
- Tensorflow implementation of proximal policy optimization (PPO) algorithm☆13Feb 28, 2018Updated 8 years ago
- Baselines and memory-based scenarios for the ViZDoom simulator☆36Dec 8, 2022Updated 3 years ago
- Research project - real-time multi-agent pursuit a moving target☆17Mar 13, 2021Updated 5 years ago
- Implementation of proximal policy optimization(PPO) with tensorflow☆35Feb 10, 2018Updated 8 years ago
- Models built with TensorFlow☆26Dec 5, 2018Updated 7 years ago
- Accompanying code for "Deep Reinforcement Learning that Matters"☆154Sep 22, 2017Updated 8 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- My homework solutions for UC Berkeley CS294: deep unsupervised learning☆14Mar 24, 2023Updated 3 years ago
- Code for the paper Novelty Search in Representational Space for Sample Efficient Exploration presented at NeurIPS 2020.☆14Jul 16, 2024Updated 2 years ago
- Normalizing Flows in Jax☆109Aug 19, 2020Updated 6 years ago
- Repository for the paper "Long-Horizon Visual Planning with Goal-Conditioned Hierarchical Predictors"☆46Nov 22, 2022Updated 3 years ago
- Code for the paper "Continuous Adaptation via Meta-Learning in Nonstationary and Competitive Environments"☆310Apr 13, 2023Updated 3 years ago
- Proximal Policy Optimization with TensorFlow and OpenAI Gym☆19Mar 31, 2018Updated 8 years ago
- self implementation of DPPO, Distributed Proximal Policy Optimization, by using tensorflow☆12Sep 1, 2017Updated 9 years ago