Deep Double Q-Learning implementation introduced by Hasselt et al in this paper: https://arxiv.org/abs/1509.06461. It's interfacing with openAI Gym. WIP.
☆31Jan 1, 2017Updated 9 years ago
Alternatives and similar repositories for DDQN
Users that are interested in DDQN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Clustering algorithms processing methods on astronomical spectra.☆10Oct 24, 2023Updated 2 years ago
- Deep Q-Networks in tensorflow☆10Apr 4, 2017Updated 9 years ago
- Double Deep Q-Learning with Prioritized Experience Replay☆35Apr 12, 2018Updated 8 years ago
- Monte Carlo Conterfactual Regret Minimization for imperfect information games☆13Mar 29, 2019Updated 7 years ago
- Deep Reinforcement Learning with Double Q-learning☆14Nov 17, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This is now the official location of the Kaldi project.☆10Aug 22, 2019Updated 6 years ago
- Simple Goal Oriented Action Planning demo written in Javascript with Phaser for studies.☆12Aug 31, 2015Updated 10 years ago
- Experiments from our work Uncertainty Quantification and Deep Ensemble☆10Nov 1, 2021Updated 4 years ago
- Implementaion of Generic L-layer Neural Network from Scratch☆12May 14, 2018Updated 8 years ago
- A multi-agent soccer simulator in a grid-world environment, with agents implementing different reinforcement learning algorithms☆13Jun 4, 2017Updated 9 years ago
- ☆10Aug 17, 2018Updated 7 years ago
- ☆14Aug 9, 2018Updated 7 years ago
- Implementation of DeDOL algorithm - Deep Reinforcement Learning based algorithm for Green Security Games with Real Time Information☆16Nov 7, 2019Updated 6 years ago
- deep reinforcement learning using demonstrations to help solve Doom environments☆10Oct 16, 2017Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Actor Critic using Kronecker-Factored Trust Region☆19Jul 3, 2018Updated 8 years ago
- General purpose, statically typed, functional programming language☆14May 30, 2026Updated last month
- Anomaly Detection Discriminative GAN (ADD-GAN)☆15Oct 9, 2017Updated 8 years ago
- Udacity Deep Reinforcement Learning Nanodegree Program☆11Jul 12, 2019Updated 7 years ago
- An interactive simulation to explain algorithmic bias.☆13Dec 3, 2022Updated 3 years ago
- A html5 football game☆17Sep 16, 2014Updated 11 years ago
- VB Diarization with Eigenvoice and HMM Priors, refactored☆14Jul 27, 2021Updated 4 years ago
- The code for NeurIPS 2020 paper: Adversarial Crowdsourcing Through Robust Rank-One Matrix Completion.☆10Oct 26, 2020Updated 5 years ago
- Trading Stock with Deep Reinforcement Learning☆24Aug 20, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆22Sep 24, 2018Updated 7 years ago
- Example code from the talk "Refactoring Python: Why and how to restructure your code" at PyCon 2016☆20May 30, 2016Updated 10 years ago
- Pagerank in Julia. An experiment in pagerank on graphs in the order of billions of edges. Currently tested with over half a billion edges…☆12Aug 14, 2013Updated 12 years ago
- ☆16Mar 7, 2019Updated 7 years ago
- Tensorflow + OpenAI Gym implementation of Deep Q-Network (DQN), Double DQN (DDQN), Dueling Network and Deep Deterministic Policy Gradient…☆80Feb 14, 2017Updated 9 years ago
- Super Mario Bros. (NES) gameplay dataset for machine learning.☆13Jul 22, 2025Updated 11 months ago
- Overlapped Speech detection in Multi-party Conversations☆22Feb 20, 2018Updated 8 years ago
- Julia bindings for AMD's clFFT library☆16Aug 22, 2023Updated 2 years ago
- Generate semi-realistic maps with Node.js and present them with leaflet☆11Dec 11, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆24Dec 13, 2018Updated 7 years ago
- Reading Group @ DMG☆11Nov 15, 2018Updated 7 years ago
- Monte carlo tree search in Go language☆32Apr 22, 2018Updated 8 years ago
- Computer Vision, 1st Project : Shape from Shading☆12Feb 24, 2014Updated 12 years ago
- Reinforced Causal Explainer for Graph Neural Networks, TPAMI2022☆42Jun 13, 2022Updated 4 years ago
- RubyGoal soccer game for Rubyists☆25Oct 12, 2017Updated 8 years ago
- reproduce some RL or Multi-Agent models☆35May 22, 2019Updated 7 years ago