Chapter 15 AlphaZero in book Deep Reinforcement Learning: code example of AlphaZero solving Gomoku game.
☆36Feb 18, 2020Updated 6 years ago
Alternatives and similar repositories for Chapter15-AlphaZero
Users that are interested in Chapter15-AlphaZero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- (Keras) Use deep Q-learning to build two Gomoku (Five-in-a-Row) agents playing against each other.☆19Oct 8, 2016Updated 9 years ago
- A Python 3 Bandit Visualization Package☆11Oct 16, 2017Updated 8 years ago
- Click Me -->☆32Mar 3, 2023Updated 3 years ago
- Trial version for prs platform (python project). Please note that the complete experience requires downloading the Unity resource.☆10Jun 26, 2024Updated 2 years ago
- Connect6 AI based on reinforcement learning☆12Sep 13, 2019Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An illustration program which visualizes the MCTS mechanism inside AlphaZero in order to provide a better understanding of how an AI make…☆19Aug 6, 2018Updated 8 years ago
- SCoRe: Training Language Models to Self-Correct via Reinforcement Learning☆16May 14, 2026Updated 2 months ago
- Compression performance of BPG, JPEG, JPEG2000 and Webp.☆12May 15, 2019Updated 7 years ago
- ☆11Sep 6, 2024Updated last year
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 6 years ago
- Implementation of Compressed SGD with Compressed Gradients in Pytorch☆13Jul 25, 2024Updated 2 years ago
- AirSim based multi uav predictive manteinance application using reinforcement learning☆26Jun 6, 2021Updated 5 years ago
- Implementation to VirtualTaobao☆13Jan 17, 2020Updated 6 years ago
- Dynamic ensemble learning based on RL and multi-objective optimization. Deep reinforcement learning and NSGA2 are combined to realize dy…☆32Jul 28, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- DQN examples codes in chapter 4☆44Mar 24, 2023Updated 3 years ago
- ADP☆13Apr 12, 2017Updated 9 years ago
- Freeplane Add-on for importing OPML files☆14Jan 24, 2017Updated 9 years ago
- iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models (ICLR2026)☆23Jun 24, 2026Updated last month
- Transfer PaddlePaddle's codes to TensorLayerX's codes☆10Feb 10, 2023Updated 3 years ago
- Common support code for user-facing front end systems.☆12Updated this week
- Google MobileNets Implementation using Tensorflow☆18Jun 6, 2017Updated 9 years ago
- ☆38May 2, 2019Updated 7 years ago
- Demo OpenCL kernels written in Zig☆13Jul 23, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A collection of free online materials for control engineering☆22Feb 4, 2025Updated last year
- Datastructures and algorithms for audio graphs☆22Aug 29, 2022Updated 3 years ago
- ☆10Dec 9, 2021Updated 4 years ago
- Marp Theme using black and white☆15Aug 2, 2024Updated 2 years ago
- Recommendation engine and it's algorithms in python , R .☆12Oct 26, 2018Updated 7 years ago
- ArXiv'18 implementation of amortized maximum likelihood (AML) for high-quality, weakly-supervised shape completion.☆11Nov 30, 2018Updated 7 years ago
- Heuristic Dynamic Programming with Python☆14Jul 28, 2014Updated 12 years ago
- papers about reinforcement learning☆13Jan 4, 2021Updated 5 years ago
- An implementation of improved AlphaGo algorithm in the game of Gomoku.☆58Nov 12, 2019Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A method adapted from the paper Nonlinear System Identification of Soft Robot Dynamics Using Koopman Operator Theory by D. Bruder et al t…☆12Sep 24, 2020Updated 5 years ago
- The figures for the Deep Learning textbook (www.deeplearningbook.org)☆17Oct 9, 2017Updated 8 years ago
- DQN with freezing target network in tensorflow on pygame FlappyBird☆11Dec 19, 2018Updated 7 years ago
- Hands-On TensorBoard for PyTorch Developers, Published by Packt☆11Dec 15, 2025Updated 7 months ago
- Model-based shared control of human-machine systems☆14Jul 26, 2018Updated 8 years ago
- Implementation of Incremental Heuristic Dynamic Programming with Neural Networks as Actor and Critic function appoximators.☆12Jan 5, 2023Updated 3 years ago
- Code for "Calibrated Model-Based Deep Reinforcement Learning", ICML 2019.☆54May 15, 2019Updated 7 years ago