OpenAI Gym Env for game Gomoku(Five-In-a-Row, 五子棋, 五目並べ, omok, Gobang,...)
☆89Oct 11, 2024Updated last year
Alternatives and similar repositories for gym-gomoku
Users that are interested in gym-gomoku are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of the AlphaZero algorithm for Gomoku (also called Gobang or Five in a Row)☆3,625Apr 24, 2024Updated 2 years ago
- Implementation of Russo and Van Roy work on Information Directed Sampling (2017)☆21Jan 18, 2019Updated 7 years ago
- ☆10May 15, 2020Updated 6 years ago
- soft q learning and soft actor critic☆16Dec 23, 2018Updated 7 years ago
- Tensorflow implementation of SNAIL and RL2☆11Aug 17, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆17Oct 21, 2023Updated 2 years ago
- a Renju game, replicate paper "Mastering the game of Go with deep neural networks and tree search"☆20Jun 29, 2016Updated 10 years ago
- Reinforcing Your Learning of Reinforcement Learning☆96Jul 14, 2019Updated 7 years ago
- A light-weight python version of moses BLEU.☆13Jan 24, 2019Updated 7 years ago
- Implementation of the AlphaZero algorithm for playing the simple board game Gomoku☆14May 22, 2023Updated 3 years ago
- SplitNet implemented based on ResNet-50 trained on ImageNet-22K☆16Jun 18, 2018Updated 8 years ago
- ☆11Jul 29, 2021Updated 5 years ago
- (Personal experiment) Unsupervised Predictive Memory in a Goal-Directed Agent https://arxiv.org/abs/1803.10760☆25May 3, 2019Updated 7 years ago
- 基于DQN的五子棋人机对弈☆62Mar 24, 2019Updated 7 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆25Oct 28, 2020Updated 5 years ago
- Julia Implementation of the POMCP algorithm for solving POMDPs☆12Aug 6, 2021Updated 5 years ago
- AI-powered cryptocurrency trading bot built using deep reinforcement learning (DRL). The bot is designed as a research platform for devel…☆11Jan 18, 2025Updated last year
- Unrailed! simulator using C++ with some reinforcement learning and Unrailed! AI using Python with OpenCV☆18Dec 6, 2021Updated 4 years ago
- Courbariaux, Matthieu, Yoshua Bengio, and Jean-Pierre David. "Binaryconnect: Training deep neural networks with binary weights during pro…☆12Aug 31, 2020Updated 6 years ago
- using python to realize GA☆22May 25, 2018Updated 8 years ago
- Implement DQN and DDQN algorithm on Atari games,such as BreakoutNoFrameskip-v4, PongNoFrameskip-v4,BoxingNoFrameskip-v4.☆15Jun 30, 2020Updated 6 years ago
- This package allows to use PLE as a gym environment.☆71Jun 26, 2020Updated 6 years ago
- Deep Deterministic Policy Gradient implemented in PyTorch for DeepMind Control Suite☆25Oct 11, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Experiments from "The Generalization-Stability Tradeoff in Neural Network Pruning": https://arxiv.org/abs/1906.03728.☆14Oct 23, 2020Updated 5 years ago
- A student implementation of Alpha Go Zero☆285Aug 1, 2018Updated 8 years ago
- 📝 A personal collection of templates for Markdown+LaTeX-based writing.☆16Oct 11, 2018Updated 7 years ago
- Implementation of the Model-Based Meta-Policy-Optimization (MB-MPO) algorithm☆45Nov 15, 2018Updated 7 years ago
- Repository for IROS 2019☆28Nov 16, 2019Updated 6 years ago
- Repository for Iterated Relearning: The Impact of Non-stationarity on Generalisation in Deep Reinforcement Learning☆11Jun 8, 2020Updated 6 years ago
- A PyTorch implementation of SVGD (Stein Variational Gradient Descent), contains all examples including bayesian inference in the paper☆12Jul 30, 2020Updated 6 years ago
- A gym game for Contra that for reinforcement learning☆10Oct 18, 2021Updated 4 years ago
- docker for UTH-BERT: https://ai-health.m.u-tokyo.ac.jp/uth-bert☆14Mar 24, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Fast graph-regularized matrix factorization☆20Oct 3, 2023Updated 2 years ago
- NumPy implementation of CNN☆14Aug 27, 2021Updated 5 years ago
- Random memory adaptation model inspired by the paper: "Memory-based parameter adaptation (MbPA)"☆24Mar 13, 2018Updated 8 years ago
- Trajectory-ranked Reward EXtrapolation (T-REX) for Inverse Reinforcement Learning - A Tensorflow implementation trained on OpenAI Gym env…☆19Jul 4, 2019Updated 7 years ago
- ☆12Sep 8, 2022Updated 3 years ago
- PyTorch implementation of Adversarial Patch☆15Jul 6, 2023Updated 3 years ago
- Here is the code for experiments related to the NIS+ framework.☆15Dec 15, 2025Updated 8 months ago