强化学习训练斗地主 / doudizhu AI using reinforcement learning.
☆20Sep 19, 2019Updated 6 years ago
Alternatives and similar repositories for doudizhu-rl
Users that are interested in doudizhu-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 基于强化学习的五子棋☆12Dec 30, 2018Updated 7 years ago
- C++/python fight the lord with pybind11 (强化学习AI斗地主), Accepted to AIIDE-2020☆164Jun 13, 2026Updated 2 months ago
- fight with landlord (斗地主AI)☆16Apr 4, 2018Updated 8 years ago
- Code accompanying the paper "TiZero: Mastering Multi-Agent Football with Curriculum Learning and Self-Play" (AAMAS 2023) 足球游戏智能体☆14May 25, 2023Updated 3 years ago
- MATLAB code for "PR2013 - A Comparative Study on Illumination Preprocessing in Face Recognition"☆13Aug 13, 2016Updated 10 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MTCNN with pycaffe☆14Nov 15, 2016Updated 9 years ago
- Pytorch Implementation of MuZero for gym environment. It support any Discrete , Box and Box2D configuration for the action space and obse…☆19Jan 24, 2023Updated 3 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- Using Generative Adversarial Networks (GANs) algorithm to detect outliers on tabular data☆13May 15, 2019Updated 7 years ago
- We implement AI for Hearthstone using open source Hearthstone simulator FirePlace: https://github.com/jleclanche/fireplace.☆11Jun 12, 2016Updated 10 years ago
- 斗地主残局破解器☆62Jan 27, 2023Updated 3 years ago
- ☆10Apr 23, 2021Updated 5 years ago
- 基于Dijkstra算法的武汉地铁路径规划☆10Jul 1, 2022Updated 4 years ago
- ☆14Dec 13, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- RC-NFQ: Regularized Convolutional Neural Fitted Q Iteration. A batch algorithm for deep reinforcement learning. Incorporates dropout regu…☆12Mar 17, 2021Updated 5 years ago
- ☆15Sep 17, 2019Updated 6 years ago
- 强化学习-中文笔记&资源-以python实例为主-由浅入深☆113Dec 1, 2020Updated 5 years ago
- A curated list of awesome Video generate resources and projects☆18Sep 27, 2024Updated last year
- Robust Reinforcement Learning Benchmark☆14Sep 22, 2024Updated last year
- Main P-Net(ECCV 2020) on Mvtec dataset☆19Dec 29, 2020Updated 5 years ago
- A minimal hackable implementation of policy gradient methods (GRPO, PPO, REINFORCE)☆17Feb 20, 2026Updated 6 months ago
- A Pygame+Pymunk Carrom Simulation Testbed for reinforcement learning. [CS747][ Foundations of Intelligent and Learning Agents]☆15Jun 24, 2019Updated 7 years ago
- Notes for papers or blog posts about ML, Robotics, CV.☆14Jan 4, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Learning how to ride a bicycle using reinforcement learning.☆13Dec 11, 2013Updated 12 years ago
- Tutorial for Pybullet☆10Sep 12, 2022Updated 3 years ago
- Zero-Shot Translation implemented by Transformer☆14Mar 24, 2023Updated 3 years ago
- 强化学习炒股,走向人生巅峰(或倾家荡产)☆55Mar 8, 2022Updated 4 years ago
- UNFINISHED: use this instead: https://github.com/JRCSTU/gearshift_calculation_tool☆15Oct 8, 2021Updated 4 years ago
- GPU Monte Carlo Tree Search with MPI☆26Jan 9, 2019Updated 7 years ago
- Lecture notes for a course on Decision and Game Theory for undergraduates studying AI☆13Dec 14, 2018Updated 7 years ago
- Code repository of GreenABR for MMSys 2022 submission☆14Apr 6, 2022Updated 4 years ago
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR'26] When Visualizing is the First Step to Reasoning: MIRA, a Benchmark for Visual Chain-of-Thought☆32Feb 14, 2026Updated 6 months ago
- A reinforcement learning algorithm for the 2048 game☆20Mar 25, 2014Updated 12 years ago
- Reinforcement learning algorithms to play Poker☆14Dec 29, 2021Updated 4 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- ☆18Dec 24, 2024Updated last year
- Deep Reinforcement Learning with Fined Grained Action Repetition☆22Jan 8, 2018Updated 8 years ago
- A small project that uses Discrete Denoising Diffusion Probabilistic Models (D3PMs), a generative model for discrete data that builds upo…☆18Aug 10, 2024Updated 2 years ago