Monte Carlo tree search in JAX, with functionality to continue search from a previous subtree
☆27Aug 9, 2026Updated 3 weeks ago
Alternatives and similar repositories for mctx-az
Users that are interested in mctx-az are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Classic MCTS example with mctx☆25May 25, 2023Updated 3 years ago
- A project that provides help for using DeepMind's mctx on gym-style environments.☆66Nov 14, 2024Updated last year
- fast + parallel AlphaZero in JAX☆112Aug 9, 2026Updated 3 weeks ago
- ☆15Jul 9, 2024Updated 2 years ago
- ☆19Jan 16, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆16May 19, 2023Updated 3 years ago
- A number of agents (PPO, MuZero) with a Perceiver-based NN architecture that can be trained to achieve goals in nethack/minihack environm…☆44Sep 19, 2022Updated 3 years ago
- Reinforcement learning with Equinox☆21Mar 4, 2025Updated last year
- Code for the simulations in the neural Kalman filtering paper☆25Jul 13, 2021Updated 5 years ago
- Pytorch Implementation of MuZero for gym environment. It support any Discrete , Box and Box2D configuration for the action space and obse…☆19Jan 24, 2023Updated 3 years ago
- Implementation and evaluation of the AXIOM architecture from the preprint "AXIOM: Learning to Play Games in Minutes with Expanding Object…☆80Jun 2, 2025Updated last year
- This is the code corresponding to our publication introducing ConvDecoder with physics-based regularization (CD+r) for MRI☆10Feb 6, 2026Updated 6 months ago
- Explorations into adversarial losses on top of autoregressive loss for language modeling☆41Dec 21, 2025Updated 8 months ago
- Applying DeepMind's MuZero algorithm to the cart pole environment in gym☆22May 6, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆17Oct 21, 2023Updated 2 years ago
- Implementation of some of the Deep Distributional Reinforcement Learning Algorithms.☆26Jun 17, 2025Updated last year
- UNet script, model, sample data☆14Feb 19, 2025Updated last year
- [NeurIPS 2023 Spotlight] LightZero: A Unified Benchmark for Monte Carlo Tree Search in General Sequential Decision Scenarios (awesome MCT…☆1,641Updated this week
- LuxCoreRender Windows Compilation Environment☆13Jun 3, 2024Updated 2 years ago
- The code for the paper "A Bayesian Approach to Online Planning" published in ICML 2024.☆13Jun 17, 2024Updated 2 years ago
- Thompson Sampling based Monte Carlo Tree Search for MDPs and POMDPs☆15Jun 20, 2016Updated 10 years ago
- ☆10May 1, 2023Updated 3 years ago
- AlphaZero for continuous control tasks☆23Dec 7, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICML 2023] Official code for "DevFormer: A Symmetric Transformer for Context-Aware Device Placement"☆23Dec 7, 2024Updated last year
- Code for the paper Alpha Zero in Continuous Action Space (A0C) (https://arxiv.org/pdf/1805.09613.pdf)☆15Jan 19, 2021Updated 5 years ago
- Personal reading list for learning-based long-horizon goal reaching methods☆17Nov 26, 2020Updated 5 years ago
- Copy all Google Fonts to a folder☆10Dec 21, 2018Updated 7 years ago
- Flexible Inference for Predictive Coding Networks in JAX.☆95Updated this week
- Pytorch Implementation of MuZero☆356Jul 23, 2023Updated 3 years ago
- Deep reinforcement learning framework for fast prototyping based on PyTorch☆14Mar 12, 2023Updated 3 years ago
- ☆11Jun 17, 2016Updated 10 years ago
- lanmt ebm☆12Jun 19, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Ludax is a domain-specific language for board games that automatically compiles into hardware-accelerated learning environments with the …☆34Aug 20, 2026Updated last week
- dcEmb, the Embecosm Dynamic Causal Modelling library☆13Aug 26, 2024Updated 2 years ago
- ⚡ Flashbax: Accelerated Replay Buffers in JAX☆281Updated this week
- The implementation of "The Kanerva Machine" with Pytorch and Pyro☆12Jun 14, 2018Updated 8 years ago
- ☆15Aug 22, 2024Updated 2 years ago
- A linguistic interface for web browsers☆20Aug 17, 2026Updated 2 weeks ago
- The official Python library for Formulaic☆18Apr 25, 2024Updated 2 years ago