fast + parallel AlphaZero in JAX
☆112Aug 9, 2026Updated last month
Alternatives and similar repositories for turbozero
Users that are interested in turbozero are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- fast + parallel AlphaZero in PyTorch☆15Jan 21, 2024Updated 2 years ago
- Monte Carlo tree search in JAX, with functionality to continue search from a previous subtree☆28Aug 9, 2026Updated last month
- ♟️ Vectorized RL game environments in JAX☆650Mar 6, 2025Updated last year
- A project that provides help for using DeepMind's mctx on gym-style environments.☆67Nov 14, 2024Updated last year
- Monte Carlo tree search in JAX☆2,666Sep 10, 2026Updated last week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Classic MCTS example with mctx☆25May 25, 2023Updated 3 years ago
- AlphaZero in JAX☆82Apr 3, 2024Updated 2 years ago
- cfrx is a collection of algorithms and tools for hardware-accelerated Counterfactual Regret Minimization (CFR) algorithms in Jax.☆40Mar 11, 2026Updated 6 months ago
- ☆26Apr 16, 2024Updated 2 years ago
- 🕹️ A diverse suite of scalable reinforcement learning environments in JAX☆863Sep 9, 2026Updated last week
- [IEEE ToG] MiniZero: An AlphaZero and MuZero Training Framework☆144Aug 18, 2026Updated last month
- Hardware-Accelerated Reinforcement Learning Algorithms in pure Jax!☆281Jun 10, 2026Updated 3 months ago
- Code for the paper "Harnessing Discrete Representations for Continual Reinforcement Learning"☆16Jun 16, 2024Updated 2 years ago
- ☆16Jul 16, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ⚡ Flashbax: Accelerated Replay Buffers in JAX☆282Aug 27, 2026Updated 3 weeks ago
- ☆55Apr 11, 2023Updated 3 years ago
- Corax: Core RL in JAX☆41Feb 22, 2024Updated 2 years ago
- Jaxplorer is a Jax reinforcement learning (RL) framework for exploring new ideas.☆12Jul 19, 2024Updated 2 years ago
- Named Tensors for Legible Deep Learning in JAX☆230Sep 1, 2026Updated 2 weeks ago
- RL Environments in JAX 🌍☆930Aug 30, 2026Updated 3 weeks ago
- Open-source codebase for EfficientZero, from "Mastering Atari Games with Limited Data" at NeurIPS 2021.☆944Dec 20, 2023Updated 2 years ago
- A dataloader, but for JAX☆20May 17, 2024Updated 2 years ago
- MuZero for Combinatorial Action Spaces: open-source codebase for MA-Gumbel-AlphaZero, MA-Sampled-AlphaZero, MA-Gumbel-MuZero and MA-Sampl…☆24Jan 22, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A standalone release of DeepMind Lab's maze generator with Python bindings.☆69Oct 3, 2023Updated 2 years ago
- Reinforcement learning in pure JAX.☆15Jun 24, 2026Updated 2 months ago
- Official implementation of the paper "On the Importance of Environments in Human-Robot Coordination", published in RSS 2021.☆16May 1, 2024Updated 2 years ago
- Minimal Decision Transformer Implementation written in Jax (Flax).☆18Aug 8, 2022Updated 4 years ago
- Distributed Reinforcement Learning accelerated by Lightning Fabric☆439Sep 14, 2026Updated last week
- (Crafter + NetHack) in JAX. ICML 2024 Spotlight.☆453Jun 20, 2026Updated 3 months ago
- ☆18Jan 16, 2025Updated last year
- Code for the ICML 2020 publication "Information Particle Filter Tree: An Online Algorithm for POMDPs with Belief-Based Rewards on Continu…☆14Jul 3, 2020Updated 6 years ago
- ☆22Apr 3, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Applying DeepMind's MuZero algorithm to the cart pole environment in gym☆22May 6, 2023Updated 3 years ago
- An implementation of DreamerV2 written in JAX, with support for running multiple random seeds of an experiment on a single GPU.☆18Jan 16, 2023Updated 3 years ago
- Ludax is a domain-specific language for board games that automatically compiles into hardware-accelerated learning environments with the …☆34Updated this week
- ☆19Apr 17, 2026Updated 5 months ago
- [ICLR 2025 Oral] OptionZero: A method for autonomously discovering and utilizing options in the MuZero algorithm☆28May 18, 2025Updated last year
- An implementation of the AlphaZero algorithm for adversarial games to be used with the machine learning framework of your choice☆11Aug 30, 2020Updated 6 years ago
- Simple single-file baselines for Q-Learning in pure-GPU setting☆245Nov 24, 2025Updated 9 months ago