基于DQN的五子棋人机对弈
☆62Mar 24, 2019Updated 7 years ago
Alternatives and similar repositories for -Reinforcement-Learning-five-in-a-row-
Users that are interested in -Reinforcement-Learning-five-in-a-row- are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A DeepQLearging AI playing Gomoku☆10Mar 3, 2019Updated 7 years ago
- 尝试了博弈树Min-Max + alpha-Beta剪枝方法,并找到了更好的适用于五子棋智能的棋局评估模型和选择模型☆54May 10, 2018Updated 8 years ago
- 用深度学习+强化学习编写的一个五子棋人工智障☆45Feb 16, 2018Updated 8 years ago
- ☆23Dec 23, 2017Updated 8 years ago
- ☆23May 3, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆14Jul 17, 2025Updated last year
- This project features a dynamic combat and traversal system inspired by Sekiro, incorporating fluid movement, precise timing, and strateg…☆14Oct 22, 2024Updated last year
- 引用整理https://blog.csdn.net/yellow_red_people/article/details/80465510 一文中PyTorch平台,利用DQN模型玩Flappy Bird游戏,是一个再励学习(强化学习)实验例子。☆53Feb 10, 2019Updated 7 years ago
- Code and Data for the paper "Evaluating Character Understanding of Large Language Models via Character Profiling from Fictional Works".☆21Jul 24, 2024Updated 2 years ago
- ☆18Oct 4, 2024Updated last year
- Released code for [sigmod'21] A learned Sketch for Subgraph Counting☆22Sep 26, 2021Updated 4 years ago
- Code release for "TempLM: Distilling Language Models into Template-Based Generators"☆14Jul 21, 2022Updated 4 years ago
- A Gobang(also known as "Five in a Row" and "Gomoku") game equipped with AlphaGo-liked AI.☆14May 1, 2020Updated 6 years ago
- Automate hyper-parameters tuning for NNs (learning rate, number of dense layers and nodes and activation function)☆14Aug 9, 2020Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- DQN stock trading pytorch implementation☆41May 31, 2026Updated 3 months ago
- 2018科大讯飞AI营销算法大赛☆20Sep 19, 2018Updated 7 years ago
- On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning☆16Apr 30, 2023Updated 3 years ago
- OpenAI Gym Env for game Gomoku(Five-In-a-Row, 五子棋, 五目並べ, omok, Gobang,...)☆89Oct 11, 2024Updated last year
- Public Repo for the paper "Overcoming The Spectral-Bias of Neural Value Approximation"☆11May 25, 2024Updated 2 years ago
- A Finite Element Approximation of a Cahn--Hilliard Tumour Model with FEniCS, by Dennis Trautwein (2020).☆11Oct 11, 2020Updated 5 years ago
- A tale of works on the complexity of first-order bilevel optimization.☆25Aug 13, 2026Updated 3 weeks ago
- 贴吧舆情监测及干预工具☆13May 10, 2017Updated 9 years ago
- Dice Scores Recognition in images and live video using CNN.☆13Dec 19, 2020Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 考研数据结构练习;目前在使用C更细致的重写:https://github.com/by777/dataStructureForC☆10Sep 26, 2018Updated 7 years ago
- Tools for geospatial analysis of radar rainfall fields☆12Nov 30, 2016Updated 9 years ago
- Code and data for "Inferring Rewards from Language in Context" [ACL 2022].☆16May 22, 2022Updated 4 years ago
- 基于Deep Qlearning Network的股票交易模型☆57May 15, 2017Updated 9 years ago
- 基于RFID的非接触式的定位与追踪☆16Jan 10, 2021Updated 5 years ago
- Generation of columnar jointed rock using Voronoi method☆11Dec 13, 2019Updated 6 years ago
- a Renju game, replicate paper "Mastering the game of Go with deep neural networks and tree search"☆20Jun 29, 2016Updated 10 years ago
- ☆11Aug 1, 2019Updated 7 years ago
- Statistical methods for estimating scaling laws in urban data☆11Dec 9, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Computer Vision Research Project☆11Aug 30, 2019Updated 7 years ago
- MaxSum is an algorithm about Distributed Constraint Optimization Problems (DCOPs)☆11Jan 15, 2018Updated 8 years ago
- Single Episode Policy Transfer in Reinforcement Learning☆17Jun 13, 2022Updated 4 years ago
- 基于强化学习的五子棋☆12Dec 30, 2018Updated 7 years ago
- 16-811 Project.☆10Jan 12, 2018Updated 8 years ago
- Flask Themes☆19Dec 7, 2022Updated 3 years ago
- Reference implementation and experiments for combining reaction-diffusion and tissue growth☆14Jan 13, 2021Updated 5 years ago