Reproducing Policy Distillation (DeepMind paper ICLR 2016)
☆22Feb 17, 2020Updated 6 years ago
Alternatives and similar repositories for policydistillation
Users that are interested in policydistillation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Model-Free-Episodic-Control implementation.☆18Jun 3, 2019Updated 7 years ago
- ☆15Nov 22, 2019Updated 6 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- Design good curriculums for deep reinforcement learning☆14May 18, 2016Updated 10 years ago
- rlcourse-march-17-hugobb created by GitHub Classroom☆15Jul 3, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code to reproduce the results in the "Unsupervised Learning of Goal Spaces for Intrinsically Motivated Exploration"☆21Feb 14, 2018Updated 8 years ago
- Core interface to design, solve, and simulate trajectory games.☆21Jul 3, 2026Updated 2 months ago
- Planning with inferred internal states of other players in general-sum differential games.☆17May 3, 2022Updated 4 years ago
- [ICLR 2020, Oral] Harnessing Structures for Value-Based Planning and Reinforcement Learning☆33Feb 1, 2020Updated 6 years ago
- Inverse Reinforcement learning proof-of-concept using the Guided Cost/Reward Learning approach☆10Mar 23, 2020Updated 6 years ago
- Multi-agent coordination using game theory and nonlinear opinion dynamics - CDC 2023☆15Nov 29, 2023Updated 2 years ago
- ☆85May 29, 2019Updated 7 years ago
- Code for Automatic Curriculum Learning through Value Disagreement☆33Jun 15, 2020Updated 6 years ago
- Implementation of Attentive Multi Task Deep Reinforcement Learning Architecture in Tensorflow☆15Apr 5, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AlphaGo Zero Reinforcement Learning Sokoban Solver☆11Jun 20, 2018Updated 8 years ago
- A reinforcement learning algorithm controller for a satellite using the orekit library☆20Feb 20, 2022Updated 4 years ago
- Project exploring Multi Task Deep Reinforcement Learning neural network architectures and algorithms with Open AI Gym and TensorFlow☆17Sep 5, 2018Updated 8 years ago
- From simulation to real world using deep generative models☆19Sep 30, 2018Updated 7 years ago
- using information theory to encourage agents to cooperate and compete☆19Oct 4, 2018Updated 7 years ago
- Control with Deep Reinforcement Learning☆16Sep 14, 2023Updated 3 years ago
- Pytorch code for Arxiv Paper: Learning to learn: Meta-Critic Networks for Sample-Efficient Learning☆57Apr 3, 2018Updated 8 years ago
- Project on Successor Features in Deep Reinforcement Learning and Transfer Learning☆25Feb 5, 2018Updated 8 years ago
- Group project "Algorithms for large-scale optimal transport". Implement ADMMs and Sinkhorn's Algorithms.☆11Jan 28, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Supporting code for "Parallel Streaming Wasserstein Barycenters"☆11Nov 14, 2017Updated 8 years ago
- Code for "Calibrated Model-Based Deep Reinforcement Learning", ICML 2019.☆55May 15, 2019Updated 7 years ago
- Code for "Demonstration-free Autonomous Reinforcement Learning via Implicit and Bidirectional Curriculum" (ICML 2023)☆10Jul 6, 2023Updated 3 years ago
- Implementation of the paper 'Stochastic Wasserstein Barycenters'☆11Oct 17, 2018Updated 7 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- The model for edge classification by transforming edges to nodes.☆15Dec 22, 2020Updated 5 years ago
- Package to support the analysis of high-precision astrometry timeseries, in particular the determination of Keplerian orbits.☆18May 28, 2025Updated last year
- Implementation of the Playground environment from the paper Language as a Cognitive Tool to Imagine Goals inCuriosity-Driven Exploration.☆11Mar 5, 2021Updated 5 years ago
- FLUIDS is a lightweight driving simulator for benchmarking Deep Reinforcement and Imitation learning algorithms.☆24May 3, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CFG-GAN: Composite functional gradient learning of generative adversarial models☆15Jul 9, 2020Updated 6 years ago
- Trust Region Policy Optimization with Generalized Advantage Estimator☆16Nov 15, 2018Updated 7 years ago
- Implementation of Population-Guided Parallel Policy Search for Reinforcement Learning☆22Jan 9, 2020Updated 6 years ago
- csl: PyTorch-based Constrained Learning☆11Jun 1, 2022Updated 4 years ago
- RVO2-3D Library Python Bindings☆22Dec 7, 2017Updated 8 years ago
- Standalone utility to encrypt files with ice encryption, that doesn't depend on Steam.☆10Aug 28, 2013Updated 13 years ago
- Gstreamer, Qt, RTSP server☆15Sep 7, 2018Updated 8 years ago