Exploration by Random Network Distillation
☆15Dec 30, 2018Updated 7 years ago
Alternatives and similar repositories for rnd
Users that are interested in rnd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The open source of FeverBasketball environment for research purpose.☆11Mar 2, 2020Updated 6 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- Model-Free-Episodic-Control implementation.☆18Jun 3, 2019Updated 7 years ago
- Learning from Trajectories via Subgoal Discovery☆12Dec 10, 2020Updated 5 years ago
- Rank TD: End-to-End Robotic Reinforcement Learning without Reward Engineering and Demonstrations☆14Oct 8, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Map-Elites based on Evolution Strategies☆34Feb 11, 2022Updated 4 years ago
- Markovian State and Action Abstractions for MDPs via Hierarchical MCTS within a POMDP Formulation☆11Jul 26, 2016Updated 10 years ago
- ☆13Nov 17, 2015Updated 10 years ago
- Implementation of Few-shot Binary Image Classification using Contrastive Learning-based Approach in PyTorch☆11May 1, 2023Updated 3 years ago
- ☆13Apr 3, 2019Updated 7 years ago
- Implementation of the skill discovery algorithm described in ICLR submission "Option Discovery using Deep Skill Chaining"☆30Sep 24, 2019Updated 7 years ago
- ForgER algorithm☆23Oct 3, 2022Updated 3 years ago
- Exploration based Reinforcement Learning. (Montezuma Revenge)☆14Jul 23, 2018Updated 8 years ago
- Surprise-based intrinsic motivation for deep reinforcement learning☆21Mar 6, 2017Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12Feb 20, 2021Updated 5 years ago
- Setup for Octo and some experiments with the model☆12Apr 11, 2024Updated 2 years ago
- Code for VIREL: A Variational Inference Framework for Reinforcement Learning☆14Dec 1, 2019Updated 6 years ago
- Tensorflow/Keras code and trained models for Episodic Curiosity Through Reachability☆205Oct 2, 2020Updated 5 years ago
- Implementation of Deepmind's Neural Episodic Control☆59May 9, 2018Updated 8 years ago
- Co-training for Policy Learning☆13Aug 8, 2019Updated 7 years ago
- The implementation of Discriminator Soft Actor Critic☆15Jan 25, 2020Updated 6 years ago
- Code for paper "Episodic Memory Deep Q-Networks" (https://arxiv.org/abs/1805.07603), IJCAI 2018☆62Sep 5, 2018Updated 8 years ago
- Hindsight policy gradients☆46Jan 31, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pytorch implementation of Planar Flow☆18Dec 2, 2019Updated 6 years ago
- Code to reproduce Supervised Policy Update (ICLR 2019)☆17Dec 8, 2022Updated 3 years ago
- Continuous Energy Minimization for Multitarget Tracking☆20Feb 9, 2022Updated 4 years ago
- ☆20Jul 14, 2020Updated 6 years ago
- Code for the Reset-free Trial and Error learning paper (RTE) experiments☆10Jan 3, 2018Updated 8 years ago
- XShell的配置导入MobaXterm☆10Jan 21, 2021Updated 5 years ago
- Algorithmic Music Composition☆11Aug 28, 2018Updated 8 years ago
- Concurrent, streaming access to the input and outputs of system processes.☆15Apr 7, 2018Updated 8 years ago
- ☆11Sep 29, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code companion of Multi-task Learning for Aggregated Data using Gaussian Processes paper☆11Apr 6, 2020Updated 6 years ago
- Related papers for offline reforcement learning (we mainly focus on representation and sequence modeling and conventional offline RL)☆19Apr 21, 2022Updated 4 years ago
- Explore the optimization landscape for direct policy learning reinforcement learning.☆52Jan 16, 2019Updated 7 years ago
- A pure-rust(with zero dependencies) fenwick tree, for the efficient computation of dynamic prefix sums.☆24Jan 4, 2026Updated 8 months ago
- Implementation of Data Efficient Reinforcement Learning in Pytorch☆20Aug 6, 2019Updated 7 years ago
- ☆17Nov 14, 2022Updated 3 years ago
- Include Supervised learning:Logistic Regression、Multilayer perceptron and Deep Convolutional Network .And Unsupervised learning:Auto Enco…☆14Jul 29, 2025Updated last year