Random Network Distillation(RND) algo in Pytorch
☆50Feb 26, 2019Updated 7 years ago
Alternatives and similar repositories for RND-Pytorch
Users that are interested in RND-Pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Random Network Distillation pytorch☆264Mar 4, 2019Updated 7 years ago
- Code for the paper "Exploration by Random Network Distillation"☆930Oct 1, 2020Updated 5 years ago
- ☆19Mar 28, 2019Updated 7 years ago
- Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)☆41Jan 28, 2019Updated 7 years ago
- This is a TensorFlow implementation of DeepMind's A Distributional Perspective on Reinforcement Learning.(C51-DDPG)☆11Sep 14, 2017Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- In Progress : State of the art Distributed Distributional Deep Deterministic Policy Gradient algorithm implementation in pytorch.☆19Jun 15, 2018Updated 8 years ago
- ☆20Apr 10, 2018Updated 8 years ago
- Forecasting library in python☆13Sep 6, 2019Updated 6 years ago
- Code to perform Model-Free Episodic Control using Aurora OPUs☆17May 19, 2020Updated 6 years ago
- N-Layered FeUdal Networks based on FeUdal Networks adapted to suit PySC2 observations☆19Sep 17, 2019Updated 6 years ago
- Random network distillation on Montezuma's Revenge and Super Mario Bros.☆56May 12, 2025Updated last year
- ☆16Jun 30, 2019Updated 7 years ago
- This is a project using Pytorch to fulfill reinforcement learning on a simple game - Gridworld☆14Jul 13, 2020Updated 6 years ago
- advantage actor-critic reinforcement learning for openai gym cartpole☆66Jul 13, 2017Updated 9 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Proximal policy optimization in PyTorch. Easy to read and understand.☆51Oct 30, 2020Updated 5 years ago
- Smart Contract and Javascript library for the Enigma Protocol☆22Apr 16, 2020Updated 6 years ago
- This repository contains implementations of the paper VUSFA☆14Mar 31, 2021Updated 5 years ago
- Pytorch implementation of "FeUdal Networks for Hierarchical Reinforcement Learning" for Montezuma's Revenge☆95Jul 27, 2022Updated 4 years ago
- DHER: Hindsight Experience Replay for Dynamic Goals (ICLR-2019)☆65Nov 8, 2019Updated 6 years ago
- This repo replicates the results Horgan et al obtained in "Distributed Prioritized Experience Replay"☆190Mar 18, 2019Updated 7 years ago
- This is the pytorch implementation of ICML 2018 paper - Self-Imitation Learning.☆67Nov 4, 2018Updated 7 years ago
- This project explores deep reinforcement learning, hybrid actor-critic approach with A3C/PPO combined with curiosity for playing Super M…☆82Jan 19, 2019Updated 7 years ago
- ☆10Apr 18, 2017Updated 9 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Pytorch implementation of Distributed Proximal Policy Optimization: https://arxiv.org/abs/1707.02286☆184Mar 25, 2018Updated 8 years ago
- Learning Transferable Features with Deep Adaptation Networks☆12Jul 18, 2023Updated 3 years ago
- Tensorflow implementation of BootstrappedDQN using OpenAI baselines☆19Jan 12, 2021Updated 5 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- Baseline for NeurIPS_Auto_Bidding_General_Track☆38Aug 8, 2024Updated 2 years ago
- An implementation of the AdaOPS (Adaptive Online Packing-based Search), which is an online POMDP Solver used to solve problems defined wi…☆16Nov 16, 2025Updated 9 months ago
- PyTorch implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Nov 22, 2019Updated 6 years ago
- TensorFlow implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Dec 8, 2022Updated 3 years ago
- Avenue is a simulator designed to test and prototype reinforcement learning algorithms. Avenue is a ServiceNow Research project that was …☆14Jul 15, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Jul 10, 2019Updated 7 years ago
- ☆11Apr 20, 2021Updated 5 years ago
- Tabu search solver for job shop scheduling problem☆17Nov 8, 2018Updated 7 years ago
- Software for performing value iteration on partially observable Markov decision processes (POMDPs).☆18Feb 2, 2024Updated 2 years ago
- A multi-task deep reinforcement learning model for trading futures contracts using the Interactive Brokers API and TensorFlow☆15Feb 8, 2023Updated 3 years ago
- Energy-Based Hindsight Experience Prioritization (CoRL 2018) Oral presentation (7%)☆35Nov 28, 2018Updated 7 years ago
- Code for the paper "A Boolean Task Algebra For Reinforcement Learning"☆10Dec 8, 2022Updated 3 years ago