强化学习训练斗地主 / doudizhu AI using reinforcement learning.
☆20Sep 19, 2019Updated 7 years ago
Alternatives and similar repositories for doudizhu-rl
Users that are interested in doudizhu-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 用强化学习来玩微信跳一跳☆21Jan 15, 2018Updated 8 years ago
- C++/python fight the lord with pybind11 (强化学习AI斗地主), Accepted to AIIDE-2020☆164Jun 13, 2026Updated 3 months ago
- Code accompanying the paper "TiZero: Mastering Multi-Agent Football with Curriculum Learning and Self-Play" (AAMAS 2023) 足球游戏智能体☆14May 25, 2023Updated 3 years ago
- Using Generative Adversarial Networks (GANs) algorithm to detect outliers on tabular data☆13May 15, 2019Updated 7 years ago
- ☆22Apr 2, 2026Updated 5 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- We implement AI for Hearthstone using open source Hearthstone simulator FirePlace: https://github.com/jleclanche/fireplace.☆11Jun 12, 2016Updated 10 years ago
- 斗地主残局破解器☆62Jan 27, 2023Updated 3 years ago
- ☆10Apr 23, 2021Updated 5 years ago
- 基于Dijkstra算法的武汉地铁路径规划☆10Jul 1, 2022Updated 4 years ago
- ☆14Dec 13, 2024Updated last year
- RC-NFQ: Regularized Convolutional Neural Fitted Q Iteration. A batch algorithm for deep reinforcement learning. Incorporates dropout regu…☆12Mar 17, 2021Updated 5 years ago
- Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow - Tensorlfow Im…☆13Feb 2, 2019Updated 7 years ago
- ☆15Sep 17, 2019Updated 7 years ago
- 强化学习-中文笔记&资源-以python实例为主-由浅入深☆114Dec 1, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A curated list of awesome Video generate resources and projects☆18Sep 27, 2024Updated last year
- 2019 Fall - Game theory and Multi-agent RL Termproject☆10Dec 13, 2019Updated 6 years ago
- Simple code for running and visualizing replicator dynamics☆11Jan 31, 2024Updated 2 years ago
- Robust Reinforcement Learning Benchmark☆14Sep 22, 2024Updated last year
- Main P-Net(ECCV 2020) on Mvtec dataset☆19Dec 29, 2020Updated 5 years ago
- An illustration program which visualizes the MCTS mechanism inside AlphaZero in order to provide a better understanding of how an AI make…☆19Aug 6, 2018Updated 8 years ago
- A minimal hackable implementation of policy gradient methods (GRPO, PPO, REINFORCE)☆18Feb 20, 2026Updated 7 months ago
- ☆22Mar 25, 2025Updated last year
- A Pygame+Pymunk Carrom Simulation Testbed for reinforcement learning. [CS747][ Foundations of Intelligent and Learning Agents]☆15Jun 24, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Notes for papers or blog posts about ML, Robotics, CV.☆14Jan 4, 2019Updated 7 years ago
- Tutorial for Pybullet☆10Sep 12, 2022Updated 4 years ago
- 强化学习炒股,走向人生巅峰(或倾家荡产)☆55Mar 8, 2022Updated 4 years ago
- UNFINISHED: use this instead: https://github.com/JRCSTU/gearshift_calculation_tool☆15Oct 8, 2021Updated 4 years ago
- Lecture notes for a course on Decision and Game Theory for undergraduates studying AI☆13Dec 14, 2018Updated 7 years ago
- 单维、多维时间序列数据预测☆11Jan 7, 2019Updated 7 years ago
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Reinforcement learning algorithms to play Poker☆14Dec 29, 2021Updated 4 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Deep Reinforcement Learning with Fined Grained Action Repetition☆22Jan 8, 2018Updated 8 years ago
- This is an unofficial implementation of ' Anomaly localization by modeling perceptual features'☆23Oct 21, 2020Updated 5 years ago
- DeepSeek R1 distilled into smaller OSS models for hobbyist☆17Dec 2, 2025Updated 9 months ago
- A small project that uses Discrete Denoising Diffusion Probabilistic Models (D3PMs), a generative model for discrete data that builds upo…☆18Aug 10, 2024Updated 2 years ago
- LTS: A DASH Streaming System for Dynamic Multi-Layer 3D Gaussian Splatting Scenes☆18Sep 4, 2025Updated last year
- Minimal example to apply Decision Transformer in Atari Pong☆16Feb 1, 2025Updated last year
- 各种环境下多智能体协同围捕算法的实现☆15Mar 28, 2021Updated 5 years ago