An implementation of AlphaZero and MCTS with neural networks for Tetris
☆22Jul 7, 2026Updated 2 months ago
Alternatives and similar repositories for alphazero-tetris
Users that are interested in alphazero-tetris are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Project explores collaboration capabilities of VDN and IQL agents on a custom MARL Food Collector environment☆11Apr 6, 2022Updated 4 years ago
- Implementation and explorations into DiscoRL, Discovering state-of-the-art reinforcement learning algorithms, David Silver's last work at…☆22Jun 13, 2026Updated 3 months ago
- An agent for playing Atari games running on a Teensy microcontroller☆14Nov 11, 2022Updated 3 years ago
- Atari-style POMDPs☆36Aug 4, 2026Updated last month
- Classic MCTS example with mctx☆25May 25, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Getting Started in Imitation Learning☆13Mar 3, 2025Updated last year
- Reinforcement Learning example in Nim, playing tic tac toe. Based off original C version from the great Antirez☆15Apr 2, 2025Updated last year
- Implementation of the VIPER algorithm introduced in "Verifiable Reinforcement Learning via Policy Extraction" by Bastani et al.☆24Nov 9, 2025Updated 10 months ago
- A blog for LLVM(v11.0.0) beginner, step by step, with detailed documents and comments. Record the way I learn LLVM.☆13Jun 17, 2022Updated 4 years ago
- Distributed constraint satisfaction with recursive message-passing agents☆16Dec 11, 2017Updated 8 years ago
- Benchmark for evaluating the generalization capabilities of Multi-Objective Reinforcement Learning (MORL) algorithms.☆31Jun 6, 2025Updated last year
- Unofficial Implementation of Null-text Inversion (https://arxiv.org/abs/2211.09794)☆12Nov 20, 2022Updated 3 years ago
- Unveiling the Layers: Neural Networks from first principles☆10Oct 1, 2025Updated 11 months ago
- ☆10Sep 21, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A Pytorch Lightning WGAN-gp to generate faces☆11Jan 26, 2021Updated 5 years ago
- A technical exploration comparing standard deep RL (PPO) against biologically plausible learning rules on a custom Pong environment. Ever…☆29May 19, 2026Updated 4 months ago
- Reading list for adversarial perspective and robustness in deep reinforcement learning.☆129Mar 2, 2026Updated 6 months ago
- Concept Learning Dynamics☆17Oct 29, 2024Updated last year
- Reading Group @mila-iqia on Computational Optimal Transport for Machine Learning Applications☆13Jun 3, 2022Updated 4 years ago
- code for the paper Imitation Learning from Observation with Automatic Discount Scheduling☆13Mar 27, 2024Updated 2 years ago
- This is a repo covers ai research papers pseudocodes☆18Jun 20, 2023Updated 3 years ago
- Code base for NeurIPS 2022 paper Curriculum Reinforcement Learning using Optimal Transport via Gradual Domain Adaptation.☆11Aug 21, 2023Updated 3 years ago
- The absolute most basic example of AlphaZero and Monte Carlo Tree Search I could come up with☆234Apr 3, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tutorial kit for building a 3D deep reinforcement learning environment with Unity ML-Agents.☆11Oct 22, 2021Updated 4 years ago
- Ray Tracer written in Rust☆13Nov 22, 2021Updated 4 years ago
- A simple but instructive implementation of DP, TP, FSDP, FSDP+TP using pytorch distributed primitives☆24Apr 12, 2026Updated 5 months ago
- A worker pool library for Rust☆13Jul 27, 2026Updated 2 months ago
- Codebase for Extracting Reward Functions from Diffusion Models☆16Dec 7, 2023Updated 2 years ago
- Develop your agent for generals.io!☆120Updated this week
- A synthetic story narration dataset to study small audio LMs.☆31Jan 21, 2024Updated 2 years ago
- An unnecessarily tiny and minimal implementation of GPT-2 in NumPy.☆11Feb 12, 2023Updated 3 years ago
- ☆18Apr 11, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- SkillHack: A Benchmark for Skill Transfer in Open-Ended Reinforcement Learning☆17Oct 23, 2022Updated 3 years ago
- a minimalistic todo app☆10May 10, 2023Updated 3 years ago
- ☆18Nov 8, 2023Updated 2 years ago
- streaming deep reinforcement learning but 4x faster with jax!☆19Jan 4, 2026Updated 8 months ago
- Gold draws from Monkey, the language featured in Thorsten Ball's books. While initially following its guidelines, Gold has (slightly) evo…☆10Jun 13, 2024Updated 2 years ago
- ☆62Mar 4, 2022Updated 4 years ago
- Official Implementation of `An Optimisation Framework for Unsupervised Environment Design` from RLC 2025☆17Nov 24, 2025Updated 10 months ago