train AI agents to master Free-style Gomoku(五子棋)
☆25Mar 2, 2024Updated 2 years ago
Alternatives and similar repositories for gomoku_rl
Users that are interested in gomoku_rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [REALM25 @ ACL25] - "StateAct" Official Paper Repo (SOTA LLM Agent)☆19Aug 7, 2026Updated last month
- ☆14May 21, 2024Updated 2 years ago
- Sim-to-real RL for in-hand cube rotation with the LEAP Hand, built on Mjlab.☆35Feb 21, 2026Updated 7 months ago
- Interactive Multi-Agent Reinforcement Learning Environment for the board game Gobblet using PettingZoo.☆12Jul 2, 2023Updated 3 years ago
- Monte Carlo Tree Search guided by neural network to play board game Gomoku (Five in a Line)☆14Jan 21, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A LLM-powered agent for NetHack☆28Nov 4, 2024Updated last year
- ☆11Nov 18, 2023Updated 2 years ago
- ☆14Feb 1, 2024Updated 2 years ago
- An implementation in python of some game agents such as AlphaBeta or MCTS, that can be applied to any n-player non deterministic game obj…☆12May 29, 2022Updated 4 years ago
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 6 months ago
- NAACL'2021: Non-Parametric Few-Shot Learning for Word Sense Disambiguation☆10Jul 1, 2021Updated 5 years ago
- A lightweight driving simulator, written in Julia.☆19Sep 25, 2024Updated 2 years ago
- Custom vim folding function☆18Jul 12, 2024Updated 2 years ago
- ☆13Oct 12, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- DreamSmooth: Improving Model-Based RL with Reward Smoothing (ICLR 2024)☆13May 6, 2024Updated 2 years ago
- Equally Spaced Stacks for SwiftUI☆21Feb 3, 2020Updated 6 years ago
- A general purpose game playing A.I. framework based on the Monte Carlo tree search algorithm.☆27Jan 4, 2023Updated 3 years ago
- Research Project on Multi-robot Target Tracking via Deep Reinforcement Learning☆21Dec 17, 2020Updated 5 years ago
- ☆14Oct 30, 2023Updated 2 years ago
- Displays the top processes according to current CPU or memory usage☆24Dec 17, 2021Updated 4 years ago
- Retarget from Human Mesh Descriptions (SMPL, SMPL-X, etc) to Humanoid Poses☆21Apr 11, 2025Updated last year
- Calculate the probability of a paper being accepted by EMNLP2023 based on score distribution of ACL2023.☆14Sep 7, 2023Updated 3 years ago
- (Keras) Use deep Q-learning to build two Gomoku (Five-in-a-Row) agents playing against each other.☆19Oct 8, 2016Updated 9 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆10Mar 11, 2024Updated 2 years ago
- An illustration program which visualizes the MCTS mechanism inside AlphaZero in order to provide a better understanding of how an AI make…☆19Aug 6, 2018Updated 8 years ago
- ☆13Jun 30, 2023Updated 3 years ago
- Code-base for the paper Spectral Normalisation for Deep Reinforcement Learning: An Optimisation Perspective.☆11Jun 26, 2021Updated 5 years ago
- Tool to perform paired evaluation of automatic systems☆13Oct 20, 2021Updated 4 years ago
- Java port of c++ version of facebook fasttext☆15Oct 14, 2019Updated 6 years ago
- Context free grammar to pushdown automaton convertor, along with string parser - Theory of Languages and Machines project, spring 2020☆13Jan 18, 2022Updated 4 years ago
- 集群算法olfati saber论文仿真☆22Dec 6, 2022Updated 3 years ago
- Contact Planning for Object Manipulation via Monte Carlo Tree Search☆15May 13, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Temporal Memory-based RRT Exploration☆27Mar 1, 2024Updated 2 years ago
- A small utility to generate JSON schemas for python functions.☆18Jun 19, 2025Updated last year
- Planning with inferred internal states of other players in general-sum differential games.☆17May 3, 2022Updated 4 years ago
- ☆15Feb 5, 2025Updated last year
- Code for experiments on transformers using Markovian data.☆22Nov 22, 2024Updated last year
- Entity linking evaluation and analysis tool☆27Apr 25, 2026Updated 5 months ago
- Code for our paper LLaMAR: LM-based Long-Horizon Planner for Multi-Agent Robotics☆41Feb 10, 2025Updated last year