Code implementation of: "Graying the black box: Understanding DQNs"
☆20Feb 23, 2017Updated 9 years ago
Alternatives and similar repositories for GrayingTheBox
Users that are interested in GrayingTheBox are comparing it to the libraries listed below
Sorting:
- Gym wrapper for pysc2☆10Sep 16, 2022Updated 3 years ago
- yet another reinforcement learning package☆12May 24, 2022Updated 3 years ago
- Meta Reinforcement Learning Experiments☆35Aug 22, 2017Updated 8 years ago
- Some time series vectorization methods which could give better representation for classification / clustering or other analysis.☆11Jan 4, 2016Updated 10 years ago
- Official implementation of "Know Your Action Set: Learning Action Relations for Reinforcement Learning", Jain et al., ICLR 2022.☆18Mar 16, 2022Updated 4 years ago
- Reinforcement Learning☆12Jun 22, 2017Updated 8 years ago
- General-purpose library for extracting interpretable models from Multi-Agent Reinforcement Learning systems☆21May 10, 2020Updated 5 years ago
- A PyTorch implementation of SSINet.☆16Nov 10, 2020Updated 5 years ago
- The Variational Homoencoder: Learning to learn high capacity generative models from few examples☆34Jul 13, 2023Updated 2 years ago
- Causal Deconvolution of Networks by Algorithmic Generative Models☆30May 9, 2019Updated 6 years ago
- Hierarchical Deep RL Network☆31Feb 20, 2017Updated 9 years ago
- Deep Reinforcement Active learning - Master Thesis☆19Dec 7, 2022Updated 3 years ago
- ☆14Dec 4, 2018Updated 7 years ago
- Code for experimenting with state and action abstractions in reinforcement learning.☆30Dec 11, 2020Updated 5 years ago
- Fully Cooperative Multi-Agent Deep Reinforcement Learning☆28Nov 20, 2019Updated 6 years ago
- 3d cartpole gym env using bullet physics trained from pixels with tensorflow LRPG, DDPG & NAF☆58Jan 2, 2017Updated 9 years ago
- Ground control station and optimization code from Tango on Quadrotors project - NTR 50759☆12Aug 26, 2018Updated 7 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 7 years ago
- JAX implementation of Graph Attention Networks☆13Jan 29, 2022Updated 4 years ago
- 🕹️ Flappy Bird hack using Deep Reinforcement Learning with Double Q-learning☆18Oct 9, 2021Updated 4 years ago
- Procgen2: A community maintained fork of procgen☆12Aug 25, 2022Updated 3 years ago
- This reposotory is for a project about Distributed TDMA for Mobile UWB Network Localization☆15Jun 1, 2021Updated 4 years ago
- A simple systemd service to better control Framework Laptop's fan *Ryzen 7040*☆15Dec 15, 2023Updated 2 years ago
- Simple GStreamer test programs for learning puporses.☆13Jul 27, 2013Updated 12 years ago
- ROMFS文件系统固件解析与提取☆12Dec 24, 2023Updated 2 years ago
- Experiments in applying interpretability techniques to learned reward functions.☆10Dec 11, 2020Updated 5 years ago
- Multilingual acoustic word embedding approaches applied and evaluated on GlobalPhone data.☆11Nov 3, 2020Updated 5 years ago
- Trust Region Policy Optimization with Generalized Advantage Estimator☆16Nov 15, 2018Updated 7 years ago
- ☆11Feb 22, 2018Updated 8 years ago
- [ICLRW'26] EoRA: Fine-tuning-free Compensation for Compressed LLM with Eigenspace Low-Rank Approximation☆29Mar 16, 2026Updated last week
- Gstreamer, Qt, RTSP server☆15Sep 7, 2018Updated 7 years ago
- Deep reinforcement learning package for torch7☆16Sep 17, 2016Updated 9 years ago
- We have a Turtlebot simulator which is treated as an autonomous vehicle. Global path planning is applied on this map environment. A webca…☆11Dec 9, 2017Updated 8 years ago
- A JAX-accelerated implementation of the Procedural Content Generation via Reinforcement Learning (PCGRL) framework. We train RL agents to…☆13Nov 26, 2025Updated 3 months ago
- This is a sample implementation of "TIMERS: Error-Bounded SVD Restart on Dynamic Networks"(AAAI 2018).☆12Jul 4, 2018Updated 7 years ago
- Companion code for Closed-Loop Koopman Operator Approximation☆16Mar 24, 2024Updated last year
- Attentional Mechanism incorporated in Asynchronous Advantage Actor Critic a3c/a2c deep mind☆10Jan 9, 2018Updated 8 years ago
- reimplementation of the ddpg algorithm using tensorflow☆38Oct 17, 2016Updated 9 years ago
- ☆11Dec 15, 2024Updated last year