AlphaZero for continuous control tasks
☆23Dec 7, 2022Updated 3 years ago
Alternatives and similar repositories for alphazero-gym
Users that are interested in alphazero-gym are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Aerial Combat environment build around PyFlyt☆12Aug 12, 2023Updated 2 years ago
- Pytorch implementation of the PAAC algorithm presented in Efficient Parallel Methods for Deep Reinforcement Learning https://arxiv.org/ab…☆20Jan 25, 2018Updated 8 years ago
- Pointax: PointMaze Environment for JAX☆28Oct 22, 2025Updated 9 months ago
- Gym environment which simulates intraday trading☆28Feb 9, 2022Updated 4 years ago
- If you want a online gym, this is the perfect page. You have some filters and inputs fields in order to find your perfect routine.☆15Mar 3, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- List of awesome JAX resources☆13Dec 8, 2022Updated 3 years ago
- Implementation of MuZero with PyTorch, based on the pseudocode from DeepMind (https://arxiv.org/src/1911.08265v2/anc/pseudocode.py).☆33Aug 14, 2022Updated 3 years ago
- Web application interface for Mathematica☆11Mar 28, 2015Updated 11 years ago
- Automatic code generator for training Reinforcement Learning policies☆11Jan 3, 2021Updated 5 years ago
- Single player Alpha Zero implementation☆42Mar 7, 2022Updated 4 years ago
- a little library to help me with things involving Koopman operators☆12Mar 3, 2022Updated 4 years ago
- Theft detector project objective is to track the surveillance area and alert the user if any movement is detected.☆18Mar 19, 2022Updated 4 years ago
- Using self-play, MCTS, and a deep neural network to create a hearthstone ai player☆31Nov 1, 2018Updated 7 years ago
- A neural network accelerated solver for mixed-strategy solutions of trajectory games. Do you even lift?☆18Jun 22, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- SIFT的代码实现 以及kmeans visual words☆11Oct 18, 2020Updated 5 years ago
- Code for "Dream and Search to Control: Latent Space Planning for Continuous Control"☆12Jul 12, 2021Updated 5 years ago
- an IP camera running your YOLO model☆11Feb 21, 2025Updated last year
- Dockerfile and instructions for building Mitsuba☆16Jun 23, 2019Updated 7 years ago
- GYM is an easy-to-use gym management and administration system. It helps you to keep track of the records of your members and their membe…☆11May 18, 2025Updated last year
- A python implementation of the COACH algorithm for the Cartpole problem in OpenAI gym.☆11Mar 15, 2019Updated 7 years ago
- ☆21Mar 5, 2023Updated 3 years ago
- ☆13Aug 23, 2023Updated 2 years ago
- Open DRUWA - Open Deep Realtime User Welcoming Assistant☆16Nov 4, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆19Jan 16, 2025Updated last year
- Jax implementation of Proximal Policy Optimization (PPO) specifically tuned for Procgen, with benchmarked results and saved model weights…☆62Aug 4, 2022Updated 3 years ago
- Implemented YOLOv2 with Tensorflow 2.0☆10Oct 6, 2022Updated 3 years ago
- C++ implementation of multi-layer feed forward neural networks with back propagation algorithm.☆10Mar 30, 2016Updated 10 years ago
- ☆12Sep 8, 2022Updated 3 years ago
- Simple implementation of V-MPO proposed in https://arxiv.org/abs/1909.12238☆48Nov 10, 2020Updated 5 years ago
- YOLO meets Optical Flow☆14Oct 13, 2022Updated 3 years ago
- Simulated Baxter Robot writing "hello"☆10Jan 15, 2016Updated 10 years ago
- ☆18Jul 10, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ICML 2022] Robust Deep Reinforcement Learning through Bootstrapped Opportunistic Curriculum☆12Jul 15, 2022Updated 4 years ago
- Algorithms for Policy Evaluation, Estimation of Action Values, Policy Improvement, Policy Iteration, Truncated Policy Evaluation, Truncat…☆11Apr 3, 2019Updated 7 years ago
- Classic MCTS example with mctx☆25May 25, 2023Updated 3 years ago
- Object detection using webcam or mobile camera in the browser. Written in Tensorflow.js☆13Apr 12, 2019Updated 7 years ago
- This repository provides a GitHub Action for running the Kani Rust Verifier in CI.☆13May 13, 2025Updated last year
- Implicit Distributional Actor Critic☆11Dec 8, 2021Updated 4 years ago
- Robot payload estimation, based on: C. G. Atkeson, C. H. An, and J. M. Hollerbach, “Estimation of Inertial Parameters of Manipulator Load…☆14Apr 13, 2022Updated 4 years ago