基于强化学习的五子棋
☆12Dec 30, 2018Updated 7 years ago
Alternatives and similar repositories for Renju
Users that are interested in Renju are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 强化学习训练斗地主 / doudizhu AI using reinforcement learning.☆20Sep 19, 2019Updated 7 years ago
- 用强化学习玩俄罗斯方块☆19Feb 4, 2018Updated 8 years ago
- 强化学习玩flappy bird☆23Feb 1, 2021Updated 5 years ago
- 用强化学习来玩微信跳一跳☆21Jan 15, 2018Updated 8 years ago
- a Renju game, replicate paper "Mastering the game of Go with deep neural networks and tree search"☆20Jun 29, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 利用Torch和强化学习训练flappy bird小游戏☆13Nov 18, 2022Updated 3 years ago
- ☆11Aug 10, 2021Updated 5 years ago
- Extension of OpenAI Gym that implements multiple two-player zero-sum 2-dimension board games☆11Sep 11, 2022Updated 4 years ago
- 用深度学习+强化学习编写的一个五子棋人工智障☆45Feb 16, 2018Updated 8 years ago
- 用深度强化学习玩合成大西瓜☆27Feb 1, 2021Updated 5 years ago
- Verkle trees with inner product argument (IPA) based polynomial commitment [Prototype]☆15Mar 26, 2022Updated 4 years ago
- AI项目(强化学习、深度学习、计算机视觉、推荐系统、自然语言处理、机器导航、医学影像处理)☆95Aug 8, 2023Updated 3 years ago
- Error detection in Knowledge Graphs: Path Ranking, Embeddings or both?☆12Jan 26, 2020Updated 6 years ago
- 本ROS教程 整理自ROS WIKI☆14Jul 3, 2018Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- An tensorflow implementation of Word2Vec model: Skip-gram, CBOW.☆16Jan 17, 2017Updated 9 years ago
- 考研数据结构练习;目前在使用C更细致的重写:https://github.com/by777/dataStructureForC☆10Sep 26, 2018Updated 8 years ago
- Implementation of Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation. Paper: https://arxiv.org/abs/2404.06809☆22Oct 22, 2024Updated last year
- 基于OpenAI Gym的程序化交易环境模拟器☆15Jul 13, 2021Updated 5 years ago
- A python implementation of PROCLUS: PROjected CLUStering algorithm.☆10Jan 12, 2015Updated 11 years ago
- Recommendation engine and it's algorithms in python , R .☆12Oct 26, 2018Updated 7 years ago
- ArXiv'18 implementation of amortized maximum likelihood (AML) for high-quality, weakly-supervised shape completion.☆11Nov 30, 2018Updated 7 years ago
- pytorch版损失函数,改写自科学空间文章,【通过互信息思想来缓解类别不平衡问题】、【将“softmax+交叉熵”推广到多标签分类问题】☆12Aug 22, 2021Updated 5 years ago
- Easy-to-use MIRAGE code for faithful answer attribution in RAG applications. Paper: https://aclanthology.org/2024.emnlp-main.347/☆25Mar 10, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Reproduction of the paper "Soft Q-Learning with Mutual Information Regularization" CoRL 2019.☆10Jan 10, 2019Updated 7 years ago
- [AAAI26] Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilitie…☆11Feb 7, 2026Updated 7 months ago
- A Python 3 Bandit Visualization Package☆11Oct 16, 2017Updated 8 years ago
- An implementation of the Jenkins Traub polynomial root finding algorithm☆14Aug 23, 2015Updated 11 years ago
- Kate-Zaverucha-Goldberg Polynomial Commitments☆29Nov 20, 2021Updated 4 years ago
- 在线直播/录播教育平台 开发中☆22Oct 13, 2020Updated 5 years ago
- Gomoku AI based AlphaZero Algorithm☆10Feb 27, 2019Updated 7 years ago
- Fast interpolative decompositions in Python☆10Jan 4, 2021Updated 5 years ago
- DXC 新一代、轻量级项目生命周期质量管理平台。DEMO:☆16Jan 23, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- uct tree search + supervised lerning for atari games☆12Feb 14, 2017Updated 9 years ago
- An attempt to apply reinforcement learning to graph signal recovery problem☆11Aug 25, 2021Updated 5 years ago
- 2018科大讯飞AI营销算法大赛模型方案☆22Oct 18, 2018Updated 7 years ago
- Interpretability dashboard for reinforcement learners☆16Jun 4, 2019Updated 7 years ago
- C#的GUI五子棋大作业 包括禁手 AI 简单直播功能☆10Dec 14, 2018Updated 7 years ago
- 机器翻译Jupyter Notebook教程☆23May 25, 2021Updated 5 years ago
- Store articles for WeChat Public 'CVDaily'☆11Feb 7, 2018Updated 8 years ago