Upside-Down Reinforcement Learning (⅂ꓤ) implementation in PyTorch. Based on the paper published by Jürgen Schmidhuber.
☆79Aug 13, 2020Updated 5 years ago
Alternatives and similar repositories for Upside-Down-Reinforcement-Learning
Users that are interested in Upside-Down-Reinforcement-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The implementation of "The Kanerva Machine" with Pytorch and Pyro☆12Jun 14, 2018Updated 8 years ago
- on-policy optimization baselines for deep reinforcement learning☆32Apr 3, 2020Updated 6 years ago
- Variational Reinforcement Learning☆18Jul 25, 2024Updated 2 years ago
- Episodic Control☆22Sep 20, 2022Updated 3 years ago
- Code for the paper Alpha Zero in Continuous Action Space (A0C) (https://arxiv.org/pdf/1805.09613.pdf)☆15Jan 19, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of "Training Agents using Upside-Down Reinforcement Learning (https://arxiv.org/pdf/1912.02877.pdf)"☆17Dec 17, 2019Updated 6 years ago
- Rainbow DQN implementation accompanying the paper "Fast and Data-Efficient Training of Rainbow" which reaches 205.7 median HNS after 10M …☆44Dec 11, 2021Updated 4 years ago
- Codes accompanying the paper "DOP: Off-Policy Multi-Agent Decomposed Policy Gradients" (ICLR 2021, https://arxiv.org/abs/2007.12322)☆51Dec 8, 2022Updated 3 years ago
- Collection of Deep Reinforcement Learning Algorithms implemented in PyTorch.☆82Oct 25, 2020Updated 5 years ago
- PhD Publications and Thesis on LASSO Model Predictive Control☆20Jun 2, 2019Updated 7 years ago
- HaVSA (Have-Saa) is a Haskell implementation of the Version Space Algebra Machine Learning technique described by Tessa Lau.☆12Jul 8, 2017Updated 9 years ago
- Model-Free Episodic Control☆14Jan 12, 2017Updated 9 years ago
- PyTorch implementation of the Munchausen Reinforcement Learning Algorithms M-DQN and M-IQN☆46Oct 4, 2020Updated 5 years ago
- Change-Based Exploration Transfer☆35Apr 24, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Jax implementation of VIT-VQGAN☆10Jan 25, 2024Updated 2 years ago
- Pytorch based implementation of Upside Down Reinforcement Learning (UDRL) by J. Schmidhuber et al.☆12May 1, 2020Updated 6 years ago
- 3D learning environment with rigid body simulation for Linux/MacOSX☆14Dec 24, 2021Updated 4 years ago
- Auxiliary variable Markov chain Monte Carlo methods☆10Oct 24, 2017Updated 8 years ago
- Open source code combining implementations of Upside Down Reinforcement Learning and Reward Conditioned Policies☆19Mar 10, 2021Updated 5 years ago
- JAX implementations of core Deep RL algorithms☆84May 2, 2022Updated 4 years ago
- MuJoCo models for Unitree Robots☆12Nov 24, 2021Updated 4 years ago
- Python-based HEX implementation for a fragment of the HEX language and a subset of features.☆15Jan 4, 2022Updated 4 years ago
- oracle-structured minimization method☆13Sep 1, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A2C training of Relational Deep Reinforcement Learning Architecture☆13Jun 22, 2022Updated 4 years ago
- ☆85Nov 19, 2020Updated 5 years ago
- Code for ICLR 2019 paper Learning Dynamics Model by Incorporating the Long Term Future☆51Jun 6, 2019Updated 7 years ago
- Generalised UDRL☆37May 12, 2022Updated 4 years ago
- DEREK (Domain Entities and Relations Extraction Kit)☆10May 22, 2023Updated 3 years ago
- Official repository for the paper "Going Beyond Linear Transformers with Recurrent Fast Weight Programmers" (NeurIPS 2021)☆52Jun 11, 2025Updated last year
- Repository for the paper "Planning to Explore via Self-Supervised World Models"☆242Feb 10, 2023Updated 3 years ago
- Code for the publication Learning to Reason with Third-Order Tensor Products.☆41Jan 14, 2019Updated 7 years ago
- Deep Q-Learning Auto Market Maker☆12Jun 12, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A python implemenation of tabular MuZero for educational purposes☆21Dec 11, 2019Updated 6 years ago
- Scaling All-Goals Updates in Reinforcement Learning Using Convolutional Neural Networks☆39Feb 5, 2020Updated 6 years ago
- A2C for GVG-AI☆22Nov 7, 2018Updated 7 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- discrete gate sizing☆14Nov 23, 2020Updated 5 years ago
- Gym environments for Robots that learn to interact with the environment autonomously☆34Dec 26, 2022Updated 3 years ago
- Source code for Pathfinding in Stochastic Environments paper.☆15Oct 27, 2022Updated 3 years ago