☆66Nov 3, 2021Updated 4 years ago
Alternatives and similar repositories for MuZeroJupyterExample
Users that are interested in MuZeroJupyterExample are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A structured implementation of MuZero☆208Jun 4, 2022Updated 4 years ago
- A simple implementation of MuZero algorithm for connect4 game☆97Aug 11, 2020Updated 6 years ago
- Pytorch Implementation of MuZero☆356Jul 23, 2023Updated 3 years ago
- AlphaGo Zero Reinforcement Learning Sokoban Solver☆11Jun 20, 2018Updated 8 years ago
- ☆28Apr 28, 2019Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Using a paper from Google DeepMind I've developed a new version of the DQN using threads exploration instead of memory replay as explain …☆84Mar 4, 2016Updated 10 years ago
- ☆18Nov 4, 2021Updated 4 years ago
- A PyTorch implementation of DeepMind's MCTSnet☆18Dec 8, 2022Updated 3 years ago
- An implementation of the AlphaZero algorithm for chess☆34Dec 8, 2022Updated 3 years ago
- soft q learning and soft actor critic☆16Dec 23, 2018Updated 7 years ago
- A C++ pytorch implementation of MuZero☆41May 18, 2026Updated 3 months ago
- ☆12Aug 15, 2020Updated 6 years ago
- Go engine with no human-provided knowledge, modeled after the AlphaGo Zero paper.☆11Jan 17, 2020Updated 6 years ago
- ☆10Apr 5, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AI for google research football☆28Dec 14, 2020Updated 5 years ago
- Chainer implementation of Self-Normalizing Networks (SNN)☆24Jun 11, 2017Updated 9 years ago
- Reinforcement Learning Assembly☆94Sep 2, 2021Updated 5 years ago
- Implement BinaryNet of CNN with chainer☆11May 5, 2016Updated 10 years ago
- Variation of "Asynchronous Methods for Deep Reinforcement Learning" with multiple processes generating experience for agent (Keras + Thea…☆44Feb 27, 2018Updated 8 years ago
- Extension of OpenAI Gym that implements multiple two-player zero-sum 2-dimension board games☆11Sep 11, 2022Updated 3 years ago
- ☆12Mar 29, 2023Updated 3 years ago
- A clean and easy implementation of MuZero, AlphaZero and Self-Play reinforcement learning algorithms for any game.☆16Oct 15, 2024Updated last year
- uct tree search + supervised lerning for atari games☆12Feb 14, 2017Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Adversarial learning by utilizing model interpretation☆10Oct 19, 2018Updated 7 years ago
- Ready to run Jupyter notebook docker image with Python 3.9, OpenCV 4 and more☆11Feb 12, 2022Updated 4 years ago
- Code for the paper "TD or not TD: Analyzing the Role of Temporal Differencing in Deep Reinforcement Learning", Artemij Amiranashvili, Ale…☆12Aug 24, 2018Updated 8 years ago
- ☆86Mar 5, 2023Updated 3 years ago
- "CoPhy: Counterfactual Learning of Physical Dynamics", F. Baradel, N. Neverova, J. Mille, G. Mori, C. Wolf, ICLR'2020☆37Apr 28, 2020Updated 6 years ago
- An example and description to Reinforcement Learning DQN model and dataformats for trading☆16Mar 30, 2019Updated 7 years ago
- Bayesian Uncertainty Exploration in Deep Reinforcement Learning☆18Jul 12, 2017Updated 9 years ago
- MuZero☆2,864Sep 3, 2024Updated last year
- Tiny语言编译器☆11Sep 2, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆10Jul 20, 2023Updated 3 years ago
- IJMLC: Open-TI: Open Traffic Intelligence with Augmented Language Model☆23Jul 30, 2025Updated last year
- Code accompanying the paper "TiZero: Mastering Multi-Agent Football with Curriculum Learning and Self-Play" (AAMAS 2023) 足球游戏智能体☆14May 25, 2023Updated 3 years ago
- Open source demo for the paper Learning to Score Behaviors for Guided Policy Optimization☆24Jun 24, 2020Updated 6 years ago
- Momentum Contrast for Unsupervised Visual Representation Learning☆16Mar 24, 2023Updated 3 years ago
- TensorFlow A2C to solve Acrobot, with synchronized parallel environments☆35Apr 21, 2018Updated 8 years ago
- Convert .vox to .obj☆14Nov 24, 2018Updated 7 years ago