Reinforcement Learning with Deep Energy-Based Policies
☆438Nov 28, 2023Updated 2 years ago
Alternatives and similar repositories for softqlearning
Users that are interested in softqlearning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Softlearning is a reinforcement learning framework for training maximum entropy policies in continuous domains. Includes the official imp…☆1,439Nov 29, 2023Updated 2 years ago
- Soft Actor-Critic☆1,307Nov 29, 2023Updated 2 years ago
- rllab is a framework for developing and evaluating reinforcement learning algorithms, fully compatible with OpenAI Gym.☆3,080Jun 10, 2023Updated 3 years ago
- ICML 2018 Self-Imitation Learning☆279Apr 18, 2020Updated 6 years ago
- ☆162Jul 21, 2017Updated 9 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Collection of reinforcement learning algorithms☆2,941Jun 17, 2024Updated 2 years ago
- Code for the paper "Curiosity-driven Exploration in Deep Reinforcement Learning via Bayesian Neural Networks"☆346Nov 22, 2018Updated 7 years ago
- Guided Policy Search☆601Feb 9, 2021Updated 5 years ago
- Reproduction of the paper "Soft Q-Learning with Mutual Information Regularization" CoRL 2019.☆10Jan 10, 2019Updated 7 years ago
- NIPS 2017 Value Prediction Network☆166Jan 12, 2018Updated 8 years ago
- Accompanying code for "Deep Reinforcement Learning that Matters"☆154Sep 22, 2017Updated 9 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago
- Implementation of TRPO and related algorithms☆654May 20, 2018Updated 8 years ago
- Code for the paper "Continuous Adaptation via Meta-Learning in Nonstationary and Competitive Environments"☆309Apr 13, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Inferring beliefs about dynamics from behavior☆30May 24, 2018Updated 8 years ago
- Implementation of 'A Distributional Perspective on Reinforcement Learning' and 'Distributional Reinforcement Learning with Quantile Regre…☆134May 5, 2019Updated 7 years ago
- Code for Stabilizing Off-Policy RL via Bootstrapping Error Reduction☆163Jul 17, 2020Updated 6 years ago
- RUDDER for ATARI games with delayed rewards in OpenAI Baselines package☆268Oct 24, 2019Updated 6 years ago
- Proximal Policy Optimization with Stein Control Variates:☆34Feb 12, 2018Updated 8 years ago
- Trust Region Policy Optimization with TensorFlow and OpenAI Gym☆364Jun 2, 2020Updated 6 years ago
- ☆346Jan 24, 2018Updated 8 years ago
- ☆273Jun 5, 2018Updated 8 years ago
- PyTorch implementation of Trust Region Policy Optimization☆450Sep 13, 2018Updated 8 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Multitask Environments for RL☆283Aug 23, 2021Updated 5 years ago
- Code for the paper "Meta-Learning Shared Hierarchies"☆619Jul 6, 2023Updated 3 years ago
- PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinfor…☆3,904May 29, 2022Updated 4 years ago
- [ICML 2017] TensorFlow code for Curiosity-driven Exploration for Deep Reinforcement Learning☆1,489Dec 7, 2022Updated 3 years ago
- Code for hierarchical imitation learning and reinforcement learning☆303Mar 14, 2018Updated 8 years ago
- Code for the paper "When to Trust Your Model: Model-Based Policy Optimization"☆562Nov 22, 2022Updated 3 years ago
- Experiment code for "Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models"☆484Jul 6, 2023Updated 3 years ago
- Code for the paper "Generative Adversarial Imitation Learning"☆726Nov 22, 2018Updated 7 years ago
- Deep Planning Network: Control from pixels by latent planning with learned dynamics☆380Oct 15, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Value Iteration Networks☆291Apr 21, 2017Updated 9 years ago
- DEPRECATED: Open-source software for robot simulation, integrated with OpenAI Gym.☆2,169Apr 2, 2023Updated 3 years ago
- Code for the paper "Emergent Complexity via Multi-agent Competition"☆836Apr 2, 2023Updated 3 years ago
- An implementation of the Augmented Random Search algorithm☆436Sep 29, 2021Updated 5 years ago
- Reinforcement learning with unsupervised auxiliary tasks☆425Feb 13, 2019Updated 7 years ago
- Stochastic Neural Networks for Hierarchical Reinforcement Learning☆93Apr 17, 2018Updated 8 years ago
- These are experiments for examining reproducibility in Policy Gradient RL algorithms in Continuous domains. Mainly using the Rllab implem…☆17Sep 20, 2017Updated 9 years ago