Code implementation of: "Graying the black box: Understanding DQNs"
☆20Feb 23, 2017Updated 9 years ago
Alternatives and similar repositories for GrayingTheBox
Users that are interested in GrayingTheBox are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- Gym wrapper for pysc2☆10Sep 16, 2022Updated 3 years ago
- Counterfactual explanations for Reinforcement Learning agents on Atari☆12Apr 3, 2023Updated 3 years ago
- yet another reinforcement learning package☆12May 24, 2022Updated 4 years ago
- Meta Reinforcement Learning Experiments☆35Aug 22, 2017Updated 9 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for our paper "Visualizing and Understanding Atari Agents" (https://goo.gl/AMAoSc)☆125Oct 21, 2021Updated 4 years ago
- This project visualizes the knowledge of an agent trained by Deep Reinforcement Learning (paper will be published) using Backpropagation,…☆18May 20, 2020Updated 6 years ago
- Reinforcement Learning☆12Jun 22, 2017Updated 9 years ago
- General-purpose library for extracting interpretable models from Multi-Agent Reinforcement Learning systems☆21May 10, 2020Updated 6 years ago
- The Variational Homoencoder: Learning to learn high capacity generative models from few examples☆34Jul 13, 2023Updated 3 years ago
- Causal Deconvolution of Networks by Algorithmic Generative Models☆30May 9, 2019Updated 7 years ago
- Hierarchical Deep RL Network☆31Feb 20, 2017Updated 9 years ago
- Deep Reinforcement Active learning - Master Thesis☆19Dec 7, 2022Updated 3 years ago
- ☆14Dec 4, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Fully Cooperative Multi-Agent Deep Reinforcement Learning☆27Nov 20, 2019Updated 6 years ago
- Code for experimenting with state and action abstractions in reinforcement learning.☆29Dec 11, 2020Updated 5 years ago
- The Easiest Pytorch Implementation of Branching-DQN☆12Feb 10, 2021Updated 5 years ago
- Random memory adaptation model inspired by the paper: "Memory-based parameter adaptation (MbPA)"☆24Mar 13, 2018Updated 8 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- 🕹️ Flappy Bird hack using Deep Reinforcement Learning with Double Q-learning☆19Oct 9, 2021Updated 4 years ago
- Brax + Pufferlib + CARBS for gpu-accelerated robotics RL☆12Jun 12, 2025Updated last year
- (Personal project) Pruning algorithm for DNNs using "lottery ticket" pruning☆10Dec 8, 2022Updated 3 years ago
- Procgen2: A community maintained fork of procgen☆12Aug 3, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Simple GStreamer test programs for learning puporses.☆13Jul 27, 2013Updated 13 years ago
- Multilingual acoustic word embedding approaches applied and evaluated on GlobalPhone data.☆11Nov 3, 2020Updated 5 years ago
- Code for "Dream and Search to Control: Latent Space Planning for Continuous Control"☆12Jul 12, 2021Updated 5 years ago
- Code associated to Automatica submission: "Equivariant Symmetries for Inertial Navigation Systems" by Alessandro Fornasier, Yixiao Ge, Pi…☆16Feb 12, 2025Updated last year
- Deep reinforcement learning package for torch7☆16Sep 17, 2016Updated 9 years ago
- Repository for code experimenting with RL and Solar Tracking☆13Apr 24, 2018Updated 8 years ago
- We have a Turtlebot simulator which is treated as an autonomous vehicle. Global path planning is applied on this map environment. A webca…☆11Dec 9, 2017Updated 8 years ago
- Trust Region Policy Optimization with Generalized Advantage Estimator☆16Nov 15, 2018Updated 7 years ago
- This is a sample implementation of "TIMERS: Error-Bounded SVD Restart on Dynamic Networks"(AAAI 2018).☆12Jul 4, 2018Updated 8 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Attentional Mechanism incorporated in Asynchronous Advantage Actor Critic a3c/a2c deep mind☆10Jan 9, 2018Updated 8 years ago
- 用DDPG/MADDPG/DQN/MADDPG+advantage实验 OpenAI开源的MPE环境☆24Jun 12, 2018Updated 8 years ago
- Gstreamer, Qt, RTSP server☆15Sep 7, 2018Updated 7 years ago
- Reinforcement Learning (RL) is believe to be a more general approach towards Artificial Intelligence (AI). RL is the foundation for many …☆13Dec 22, 2022Updated 3 years ago
- ☆12Dec 15, 2024Updated last year
- Gym environments for Robots that learn to interact with the environment autonomously☆34Dec 26, 2022Updated 3 years ago
- ☆16Jun 17, 2024Updated 2 years ago