Single player Alpha Zero implementation
☆42Mar 7, 2022Updated 4 years ago
Alternatives and similar repositories for alphazero_singleplayer
Users that are interested in alphazero_singleplayer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆17Oct 21, 2023Updated 2 years ago
- Applying DeepMind's MuZero algorithm to the cart pole environment in gym☆22May 6, 2023Updated 3 years ago
- fast + parallel AlphaZero in PyTorch☆15Jan 21, 2024Updated 2 years ago
- AlphaZero for continuous control tasks☆23Dec 7, 2022Updated 3 years ago
- Automatic Reparameterisation of Probabilistic Programs☆36Jun 11, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- An illustration program which visualizes the MCTS mechanism inside AlphaZero in order to provide a better understanding of how an AI make…☆19Aug 6, 2018Updated 8 years ago
- ☆13Updated this week
- ☆18May 17, 2019Updated 7 years ago
- Pytorch Implementation of MuZero for gym environment. It support any Discrete , Box and Box2D configuration for the action space and obse…☆19Jan 24, 2023Updated 3 years ago
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- Code for the Hamiltonian Variational Auto-Encoder from the proceedings of NeurIPS 2018☆16Oct 2, 2019Updated 6 years ago
- Single Player Monte Carlo Tree Search implementation☆19Jan 22, 2020Updated 6 years ago
- Monte Carlo Tree Search for Markov decision processes using the POMDPs.jl framework☆81Nov 16, 2025Updated 9 months ago
- ☆15Jun 10, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An AI agent that use Double Deep Q-learning to teach itself to land a Lunar Lander on OpenAI universe☆17Mar 15, 2021Updated 5 years ago
- Thompson Sampling based Monte Carlo Tree Search for MDPs and POMDPs☆15Jun 20, 2016Updated 10 years ago
- ☆12Apr 17, 2023Updated 3 years ago
- Capstone Research Project in NYU Courant☆12Jan 3, 2020Updated 6 years ago
- This program is used for solving Poisson Equation with several methods. And each methods are parallelized with openMP, MPI and GPU☆12Oct 25, 2017Updated 8 years ago
- Source code for "Taming GANs with Lookahead–Minmax", ICLR 2021.☆15Mar 28, 2021Updated 5 years ago
- Reward Propagation using Graph Convolutional Networks☆13Jun 19, 2021Updated 5 years ago
- ☆10Sep 23, 2021Updated 4 years ago
- Official gym API for game FightingICE.☆12Jun 27, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆27Feb 24, 2024Updated 2 years ago
- ☆14Jul 21, 2022Updated 4 years ago
- Code for the paper Novelty Search in Representational Space for Sample Efficient Exploration presented at NeurIPS 2020.☆14Jul 16, 2024Updated 2 years ago
- THEANO-KALDI-RNNs is a project implementing various Recurrent Neural Networks (RNNs) for RNN-HMM speech recognition. The Theano Code is c…☆35Apr 15, 2018Updated 8 years ago
- env for gym, match3 game☆11Jun 2, 2019Updated 7 years ago
- PyTorch implementation for our NeurIPS 2023 spotlight paper "Let the Flows Tell: Solving Graph Combinatorial Optimization Problems with G…☆69May 30, 2023Updated 3 years ago
- iOS SDK for app integrations☆16Updated this week
- Code release for the paper "Goal Representations for Instruction Following: A Semi-Supervised Language Interface to Control"☆17Apr 9, 2024Updated 2 years ago
- See https://youtube-dl.org/☆10Oct 24, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Python scripts to facilitate easy working☆11Mar 23, 2026Updated 5 months ago
- Python library to control GX11(Dexterous Hand) and EX12(Exoskeleton Glove)☆17Aug 30, 2025Updated last year
- JPEG-LM: LLMs as Image Generators with Canonical Codec Representations☆16Sep 29, 2024Updated last year
- Learning Laplacian Representations in Reinforcement Learning☆18Jan 2, 2021Updated 5 years ago
- Reinforcement learning modular with pytorch☆11Jan 18, 2021Updated 5 years ago
- ☆15Aug 14, 2026Updated 3 weeks ago
- A clean implementation based on AlphaZero for any game in any framework + tutorial + Othello/Gobang/TicTacToe/Connect4 and more☆4,514Jan 1, 2025Updated last year