Deep learning implementations (Asynchronous Deep Q-Learning) of multiple Game Theory algorithms for adversarial learning (WoLF-PHC, GIGA-WoLF, WPL, EMA-QL, PGA-APP)
☆15Sep 19, 2017Updated 8 years ago
Alternatives and similar repositories for Mixed-Policy-Asynchronous-Deep-Q-Learning
Users that are interested in Mixed-Policy-Asynchronous-Deep-Q-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Using WoLF (win or learn fast) PHC (policy hill climbing) algorithm to implement stochastic games☆15Jun 14, 2019Updated 7 years ago
- Multi-agent reinforcement learning programs based on Game theory☆42Feb 11, 2023Updated 3 years ago
- 强化学习中纳什Qlearning 实现矩阵博弈☆31Feb 25, 2019Updated 7 years ago
- Testing different RL algorithms for multi-agent environments. From SARSA, QLearning to Independent Q-Learning, Joint Action Learning and …☆12Mar 29, 2019Updated 7 years ago
- Source code for journal paper "Multiagent Reinforcement Learning With Sparse Interactions by Negotiation and Knowledge Transfer"☆13Dec 26, 2017Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This is the code provided to support the paper of "Stackelberg Game Theory Based Optimization Model for the Design of Payment Mechanism i…☆27Jul 26, 2019Updated 7 years ago
- 2019 Fall - Game theory and Multi-agent RL Termproject☆10Dec 13, 2019Updated 6 years ago
- Combination of RL and game theory for modeling human behavior☆11Aug 27, 2021Updated 4 years ago
- Exploring the Dyna-Q reinforcement learning algorithm☆17Feb 27, 2018Updated 8 years ago
- Distributed Tensorflow Implementation of Asynchronous Methods for Deep Reinforcement Learning☆28Dec 26, 2017Updated 8 years ago
- 基于RLCard平台的麻将mahjong博弈游戏代码,包括基于规则和基于Dueling DQN的Agent模型。☆32Apr 25, 2022Updated 4 years ago
- Counterfactual Regret Minimization (CFR) sample code in Python☆14Apr 16, 2019Updated 7 years ago
- Neural Fictitious Self-Play in Leduc Holdem☆11Jul 4, 2018Updated 8 years ago
- Python package for computing partial information decomposition.☆13Mar 15, 2019Updated 7 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆28Jun 3, 2022Updated 4 years ago
- Paper list of federated learning: About system design☆13Apr 13, 2022Updated 4 years ago
- A simulator for User Equipment association in UAV-empowered Emergency Networks based on Game Theory and developed in the context of Compu…☆17Apr 9, 2021Updated 5 years ago
- PyTorch implementation of Count-Based Exploration with Neural Density Models☆10Mar 22, 2018Updated 8 years ago
- yet another reinforcement learning package☆12May 24, 2022Updated 4 years ago
- Code for Optimistic Exploration even with a Pessimistic Initialisation☆14Aug 4, 2020Updated 6 years ago
- A StarCraft 2 agent for harvesting resources☆13Jun 12, 2018Updated 8 years ago
- This work shows the viability of automatically generated attack graphs that are used for adversary behavior execution in industrial contr…☆12Jun 3, 2021Updated 5 years ago
- 本科毕业设计:《多智能体博弈兵棋推演理论与验证平台设计》的源代码附录内容。强化学习算法的实现上参考了周沫凡先生的开源代码https://github.com/MorvanZhou/Reinforcement-learning-with-tensorflow☆63Jun 10, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 用户埋点行为日志分析平台,项目主要用于搭建基于Flink、Apache Doris、Redis和MySQL等中间件的用户行为日志收集、存储、分析平台,支持用户自定义查询条件☆11Dec 28, 2023Updated 2 years ago
- original source code of the ASE 2019 paper: Wuji: Automatic Online Combat Game Testing Using Evolutionary Deep Reinforcement Learning☆28Jun 8, 2020Updated 6 years ago
- Python Simulator for Evolutionary Game Theory on Networks☆20Sep 19, 2018Updated 7 years ago
- Using gbdt+lr in recommend system and comparing the auc of lr, gbdt, gbdt+lr.☆23Jul 15, 2017Updated 9 years ago
- A P2P network security monitoring system for the Ethereum blockchain.☆16Aug 7, 2024Updated 2 years ago
- Energy-Efficient Power and Subcarrier Allocation for OFDMA Systems with Value Function Approximation Approach. EI paper from march to sep…☆13Mar 13, 2017Updated 9 years ago
- The AI Arena: A framework for distributed multi-agent reinforcement learning☆14Aug 5, 2022Updated 4 years ago
- Multi agent simulation code for 2×2 Game on complex network☆30Jul 28, 2022Updated 4 years ago
- [DEPRECATED] Simulation Framework for Virtual Machine Placement in Cloud Computing Environments. [CURRENT]: https://github.com/SDDCVMP/VM…☆11May 20, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Survey of hyperparameter optimization use in NIPS2014☆12Sep 29, 2015Updated 10 years ago
- Decoupling Dynamics and Reward for Transfer Learning☆16Sep 7, 2018Updated 7 years ago
- Implementation of the Lemke-Howson algorithm for finding MNE☆15Nov 2, 2013Updated 12 years ago
- Deep Reinforcement Learning framework based on TensorFlow and OpenAI Gym☆13Apr 30, 2018Updated 8 years ago
- Learning Individual Intrinsic Reward in MARL☆65Dec 8, 2022Updated 3 years ago
- Target localization and pursuit with multiple robots☆29May 20, 2023Updated 3 years ago
- Othello reinforcement learning game playing engine☆16Jun 30, 2016Updated 10 years ago