Random Network Distillation(RND) algo in Pytorch
☆50Feb 26, 2019Updated 7 years ago
Alternatives and similar repositories for RND-Pytorch
Users that are interested in RND-Pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Random Network Distillation pytorch☆263Mar 4, 2019Updated 7 years ago
- Code for the paper "Exploration by Random Network Distillation"☆930Oct 1, 2020Updated 5 years ago
- ☆19Mar 28, 2019Updated 7 years ago
- Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)☆41Jan 28, 2019Updated 7 years ago
- This is a TensorFlow implementation of DeepMind's A Distributional Perspective on Reinforcement Learning.(C51-DDPG)☆11Sep 14, 2017Updated 8 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- In Progress : State of the art Distributed Distributional Deep Deterministic Policy Gradient algorithm implementation in pytorch.☆19Jun 15, 2018Updated 8 years ago
- ☆20Apr 10, 2018Updated 8 years ago
- Forecasting library in python☆13Sep 6, 2019Updated 6 years ago
- Implementation of PPO in Pytorch☆41Dec 6, 2017Updated 8 years ago
- N-Layered FeUdal Networks based on FeUdal Networks adapted to suit PySC2 observations☆19Sep 17, 2019Updated 6 years ago
- Random network distillation on Montezuma's Revenge and Super Mario Bros.☆55May 12, 2025Updated last year
- ☆16Jun 30, 2019Updated 7 years ago
- Proximal policy optimization in PyTorch. Easy to read and understand.☆51Oct 30, 2020Updated 5 years ago
- This repository contains implementations of the paper VUSFA☆14Mar 31, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Pytorch implementation of "FeUdal Networks for Hierarchical Reinforcement Learning" for Montezuma's Revenge☆95Jul 27, 2022Updated 3 years ago
- PyTorch Implementation of Visual GAIL in Atari Games☆14Dec 7, 2022Updated 3 years ago
- This repo replicates the results Horgan et al obtained in "Distributed Prioritized Experience Replay"☆190Mar 18, 2019Updated 7 years ago
- multi-task learning for active fault-tolerant control in quadruped robots☆14Aug 31, 2024Updated last year
- This is the pytorch implementation of ICML 2018 paper - Self-Imitation Learning.☆67Nov 4, 2018Updated 7 years ago
- Implementation prototype of the Deep Deterministic Off-Policy Gradient (DD-OPG) method.☆11Jun 12, 2019Updated 7 years ago
- This project explores deep reinforcement learning, hybrid actor-critic approach with A3C/PPO combined with curiosity for playing Super M…☆82Jan 19, 2019Updated 7 years ago
- Exploring bayesian strategies for approximating optimal actions in POMDPs☆14Jun 27, 2019Updated 7 years ago
- Pytorch implementation of Distributed Proximal Policy Optimization: https://arxiv.org/abs/1707.02286☆184Mar 25, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- PyTorch implementation of R2D2 (Recurrent Replay Distributed DPG (not DQN))☆14Mar 22, 2019Updated 7 years ago
- Tensorflow implementation of BootstrappedDQN using OpenAI baselines☆19Jan 12, 2021Updated 5 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- An implementation of the AdaOPS (Adaptive Online Packing-based Search), which is an online POMDP Solver used to solve problems defined wi…☆16Nov 16, 2025Updated 8 months ago
- ☆10Jan 21, 2021Updated 5 years ago
- PyTorch implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Nov 22, 2019Updated 6 years ago
- ☆15Jul 10, 2019Updated 7 years ago
- ☆11Apr 20, 2021Updated 5 years ago
- This is pytorch version of maddpg.☆10Jun 23, 2020Updated 6 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Tabu search solver for job shop scheduling problem☆17Nov 8, 2018Updated 7 years ago
- Software for performing value iteration on partially observable Markov decision processes (POMDPs).☆17Feb 2, 2024Updated 2 years ago
- Code for Policy Consolidation for Continual Reinforcement Learning☆10May 12, 2019Updated 7 years ago
- Sumo OSM short usage tutorial☆15Feb 7, 2018Updated 8 years ago
- Energy-Based Hindsight Experience Prioritization (CoRL 2018) Oral presentation (7%)☆35Nov 28, 2018Updated 7 years ago
- Code for the paper "A Boolean Task Algebra For Reinforcement Learning"☆11Dec 8, 2022Updated 3 years ago
- Dataset2024☆12Jun 12, 2025Updated last year