强化学习训练斗地主 / doudizhu AI using reinforcement learning.
☆20Sep 19, 2019Updated 6 years ago
Alternatives and similar repositories for doudizhu-rl
Users that are interested in doudizhu-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用强化学习来玩微信跳一跳☆21Jan 15, 2018Updated 8 years ago
- This repository is the accompanying code for the paper CFVFP. This paper presents a new algorithm for solving incomplete information game…☆16Feb 23, 2025Updated last year
- Code accompanying the paper "TiZero: Mastering Multi-Agent Football with Curriculum Learning and Self-Play" (AAMAS 2023) 足球游戏智能体☆14May 25, 2023Updated 3 years ago
- MATLAB code for "PR2013 - A Comparative Study on Illumination Preprocessing in Face Recognition"☆13Aug 13, 2016Updated 9 years ago
- MTCNN with pycaffe☆14Nov 15, 2016Updated 9 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Pytorch Implementation of MuZero for gym environment. It support any Discrete , Box and Box2D configuration for the action space and obse…☆19Jan 24, 2023Updated 3 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- Torch implementation of Multi-digit Number Recognition from Street View Imagery using Deep Convolutional Neural Networks (http://arxiv.or…☆15Jun 3, 2016Updated 10 years ago
- Unofficial Supplementary Materials for Reinforcement Learning Course at CUHK: textbooks, slides, related papers, assignment, code ...☆28Oct 27, 2020Updated 5 years ago
- Using Generative Adversarial Networks (GANs) algorithm to detect outliers on tabular data☆13May 15, 2019Updated 7 years ago
- ☆10Mar 6, 2023Updated 3 years ago
- 3rd-place solution to the ARC-AGI-3 Preview Challenge — Explore It Till You Solve It☆20Jan 7, 2026Updated 7 months ago
- We implement AI for Hearthstone using open source Hearthstone simulator FirePlace: https://github.com/jleclanche/fireplace.☆11Jun 12, 2016Updated 10 years ago
- Robust Reinforcement Learning Benchmark☆13Sep 22, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 斗地主残局破解器☆62Jan 27, 2023Updated 3 years ago
- ☆10Apr 23, 2021Updated 5 years ago
- 基于Dijkstra算法的武汉地铁路径规划☆10Jul 1, 2022Updated 4 years ago
- ☆14Dec 13, 2024Updated last year
- RC-NFQ: Regularized Convolutional Neural Fitted Q Iteration. A batch algorithm for deep reinforcement learning. Incorporates dropout regu…☆12Mar 17, 2021Updated 5 years ago
- Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow - Tensorlfow Im…☆13Feb 2, 2019Updated 7 years ago
- 强化学习-中文笔记&资源-以python实例为主-由浅入深☆112Dec 1, 2020Updated 5 years ago
- 2019 Fall - Game theory and Multi-agent RL Termproject☆10Dec 13, 2019Updated 6 years ago
- Simple code for running and visualizing replicator dynamics☆11Jan 31, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An illustration program which visualizes the MCTS mechanism inside AlphaZero in order to provide a better understanding of how an AI make…☆19Aug 6, 2018Updated 8 years ago
- A minimal hackable implementation of policy gradient methods (GRPO, PPO, REINFORCE)☆17Feb 20, 2026Updated 5 months ago
- A Pygame+Pymunk Carrom Simulation Testbed for reinforcement learning. [CS747][ Foundations of Intelligent and Learning Agents]☆15Jun 24, 2019Updated 7 years ago
- Zero-Shot Translation implemented by Transformer☆14Mar 24, 2023Updated 3 years ago
- 强化学习炒股,走向人生巅峰(或倾家荡产)☆56Mar 8, 2022Updated 4 years ago
- UNFINISHED: use this instead: https://github.com/JRCSTU/gearshift_calculation_tool☆15Oct 8, 2021Updated 4 years ago
- GPU Monte Carlo Tree Search with MPI☆26Jan 9, 2019Updated 7 years ago
- 🕵 Code for our EMNLP 2025 Main paper: "FlashAdventure: A Benchmark for GUI Agents Solving Full Story Arcs in Diverse Adventure Games"☆27Apr 26, 2026Updated 3 months ago
- Lecture notes for a course on Decision and Game Theory for undergraduates studying AI☆13Dec 14, 2018Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code repository of GreenABR for MMSys 2022 submission☆14Apr 6, 2022Updated 4 years ago
- recaptcha with lstm and mxnet☆28Mar 3, 2017Updated 9 years ago
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- 使用强化学习算法Q-learning,对3D打印的路径进行规划,减少打印喷头转弯、启停,提高打印效率。☆13Jun 30, 2021Updated 5 years ago
- [CVPR'26] When Visualizing is the First Step to Reasoning: MIRA, a Benchmark for Visual Chain-of-Thought☆32Feb 14, 2026Updated 5 months ago
- Reinforcement learning algorithms to play Poker☆14Dec 29, 2021Updated 4 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year