基于DQN的五子棋人机对弈
☆62Mar 24, 2019Updated 7 years ago
Alternatives and similar repositories for -Reinforcement-Learning-five-in-a-row-
Users that are interested in -Reinforcement-Learning-five-in-a-row- are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A DeepQLearging AI playing Gomoku☆10Mar 3, 2019Updated 7 years ago
- 通过python3.6编程,利用DQN算法实现机器学习避开障碍走到迷宫终点。(Through python3.6 programming, I use DQN algorithm to achieve machine learning and avoid obstacles…☆10Apr 15, 2018Updated 8 years ago
- 尝试了博弈树Min-Max + alpha-Beta剪枝方法,并找到了更好的适用于五子棋智能的棋局评估模型和选择模型☆54May 10, 2018Updated 8 years ago
- 用深度学习+强化学习编写的一个五子棋人工智障☆45Feb 16, 2018Updated 8 years ago
- 基于博弈树α-β剪枝搜索的五子棋AI☆789Jul 14, 2017Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 用Python语言和Tkinter图形库实现的一个简单的五子棋程序☆18Jun 7, 2016Updated 10 years ago
- ☆22May 3, 2025Updated last year
- ♟♟♟♟♟ A Gomoku game AI based on Monte Carlo Tree Search, can be trained on policy-value network now. 一个蒙特卡洛树搜索算法实现的五子棋 AI,现可用神经网络训练模型。☆51Apr 10, 2020Updated 6 years ago
- 引用整理https://blog.csdn.net/yellow_red_people/article/details/80465510 一文中PyTorch平台,利用DQN模型玩Flappy Bird游戏,是一个再励学习(强化学习)实验例子。☆53Feb 10, 2019Updated 7 years ago
- Code used for the master thesis at MIIS (UPF)☆16Dec 1, 2016Updated 9 years ago
- Use puppeteer driven Headless Chrome to generate images for arbitary HTML☆11Sep 4, 2020Updated 5 years ago
- Code and Data for the paper "Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works".☆21Jul 24, 2024Updated 2 years ago
- D3QN 强化学习打只狼☆30Feb 15, 2022Updated 4 years ago
- This project solves self-made maze in a variety of ways: A-star, Q-learning and Deep Q-network.☆28Apr 1, 2017Updated 9 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆30May 24, 2025Updated last year
- A Read-time MIDI visualization tool using PyQt☆10Jul 22, 2026Updated last week
- DQN stock trading pytorch implementation☆41May 31, 2026Updated last month
- OpenAI Gym Env for game Gomoku(Five-In-a-Row, 五子棋, 五目並べ, omok, Gobang,...)☆89Oct 11, 2024Updated last year
- ☆12Aug 28, 2020Updated 5 years ago
- 贴吧舆情监测及干预工具☆13May 10, 2017Updated 9 years ago
- Dice Scores Recognition in images and live video using CNN.☆13Dec 19, 2020Updated 5 years ago
- Tools for geospatial analysis of radar rainfall fields☆12Nov 30, 2016Updated 9 years ago
- 考研数据结构练习;目前在使用C更细致的重写:https://github.com/by777/dataStructureForC☆10Sep 26, 2018Updated 7 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Predicting 2D Steady State Fluid Flow Fields using Convolutional Neural Networks☆12Oct 3, 2020Updated 5 years ago
- Desktop Debugger for CS303 (Artificial Intelligence) Gomoku Project / 和自己的五子棋 AI 桌面对战☆16Sep 27, 2019Updated 6 years ago
- Rotation-Stacked Visibility Graph☆22Apr 21, 2026Updated 3 months ago
- Final Thesis at Fudan University, built a trading strategy on Bitcoin market using recurrent reinforcement learning☆27Nov 5, 2018Updated 7 years ago
- 基于RFID的非接触式的定位与追踪☆16Jan 10, 2021Updated 5 years ago
- ☆11Aug 1, 2019Updated 6 years ago
- a Renju game, replicate paper "Mastering the game of Go with deep neural networks and tree search"☆20Jun 29, 2016Updated 10 years ago
- Statistical methods for estimating scaling laws in urban data☆11Dec 9, 2024Updated last year
- Computer Vision Research Project☆11Aug 30, 2019Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ibox-wtoken-server☆22Jul 4, 2022Updated 4 years ago
- Matlab version of all the code of Lorena A. Barba's 12 steps to Navier stokes☆13Jul 17, 2019Updated 7 years ago
- MaxSum is an algorithm about Distributed Constraint Optimization Problems (DCOPs)☆11Jan 15, 2018Updated 8 years ago
- Single Episode Policy Transfer in Reinforcement Learning☆17Jun 13, 2022Updated 4 years ago
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆16May 19, 2023Updated 3 years ago
- Multi-Candidate Speculative Decoding☆41Apr 22, 2024Updated 2 years ago
- 基于强化学习的五子棋☆12Dec 30, 2018Updated 7 years ago