☆66Nov 3, 2021Updated 4 years ago
Alternatives and similar repositories for MuZeroJupyterExample
Users that are interested in MuZeroJupyterExample are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Example implementation of Alpha Zero' s algotirhm on Jupyter notebook☆15Nov 21, 2019Updated 6 years ago
- A python implemenation of tabular MuZero for educational purposes☆21Dec 11, 2019Updated 6 years ago
- A structured implementation of MuZero☆206Jun 4, 2022Updated 4 years ago
- A simple implementation of MuZero algorithm for connect4 game☆96Aug 11, 2020Updated 6 years ago
- Pytorch Implementation of MuZero☆356Jul 23, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- AlphaGo Zero Reinforcement Learning Sokoban Solver☆11Jun 20, 2018Updated 8 years ago
- Tabula Rasa Tic-Tac-Toe☆10Jan 3, 2019Updated 7 years ago
- ☆28Apr 28, 2019Updated 7 years ago
- Using a paper from Google DeepMind I've developed a new version of the DQN using threads exploration instead of memory replay as explain …☆84Mar 4, 2016Updated 10 years ago
- A PyTorch implementation of DeepMind's MCTSnet☆18Dec 8, 2022Updated 3 years ago
- Puzzle generator for chess variants☆19Jul 1, 2026Updated last month
- soft q learning and soft actor critic☆16Dec 23, 2018Updated 7 years ago
- ☆18Apr 19, 2024Updated 2 years ago
- The Fast Lines of Code Counter☆11Aug 22, 2025Updated 11 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Server side code of the Leela Zero project☆66Dec 8, 2022Updated 3 years ago
- The architecture used to train the level generator in the game Relay.☆12Apr 8, 2017Updated 9 years ago
- Go engine with no human-provided knowledge, modeled after the AlphaGo Zero paper.☆11Jan 17, 2020Updated 6 years ago
- ☆10Apr 5, 2024Updated 2 years ago
- AI for google research football☆28Dec 14, 2020Updated 5 years ago
- Code for Expert Supervised Reinforcement Learning☆10Apr 7, 2021Updated 5 years ago
- Implement BinaryNet of CNN with chainer☆11May 5, 2016Updated 10 years ago
- Variation of "Asynchronous Methods for Deep Reinforcement Learning" with multiple processes generating experience for agent (Keras + Thea…☆44Feb 27, 2018Updated 8 years ago
- This code illustrates the use of genetic programming to evolve financial trading strategies for a single equity stock. Individuals (strat…☆25Feb 24, 2019Updated 7 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆12Mar 29, 2023Updated 3 years ago
- Minimal code for A Generalist Agent☆44Nov 4, 2022Updated 3 years ago
- uct tree search + supervised lerning for atari games☆12Feb 14, 2017Updated 9 years ago
- Using a modified version of Werner Duvaud's MuZero implementation (https://github.com/werner-duvaud/muzero-general) this reinforcement ag…☆19Jun 24, 2026Updated last month
- ☆10Sep 20, 2018Updated 7 years ago
- A tutorial for using Hadoop with Python and Hive☆10May 26, 2015Updated 11 years ago
- Ready to run Jupyter notebook docker image with Python 3.9, OpenCV 4 and more☆11Feb 12, 2022Updated 4 years ago
- Code for the paper "TD or not TD: Analyzing the Role of Temporal Differencing in Deep Reinforcement Learning", Artemij Amiranashvili, Ale…☆12Aug 24, 2018Updated 7 years ago
- ☆84Mar 5, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- The code for Template-GPT-2 Generation Model for Logic2Text Dataset☆18Jun 1, 2020Updated 6 years ago
- An example and description to Reinforcement Learning DQN model and dataformats for trading☆16Mar 30, 2019Updated 7 years ago
- MuZero☆2,857Sep 3, 2024Updated last year
- Use tensorflow2 achieve PPO to play atari game☆13Oct 25, 2019Updated 6 years ago
- Open source demo for the paper Learning to Score Behaviors for Guided Policy Optimization☆24Jun 24, 2020Updated 6 years ago
- Momentum Contrast for Unsupervised Visual Representation Learning☆16Mar 24, 2023Updated 3 years ago
- A chess adaption of GCP's Leela Zero☆14Jan 9, 2018Updated 8 years ago