☆17Dec 4, 2019Updated 6 years ago
Alternatives and similar repositories for QMIX-Starcraft
Users that are interested in QMIX-Starcraft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 2019 Fall - Game theory and Multi-agent RL Termproject☆10Dec 13, 2019Updated 6 years ago
- qmix☆23May 28, 2020Updated 6 years ago
- Improving upon state of the art cooperative deep reinforcement learning in StarCraft II☆13May 16, 2019Updated 7 years ago
- ☆26Apr 12, 2018Updated 8 years ago
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The project to learn the QMIX.☆13Dec 19, 2019Updated 6 years ago
- ☆10Sep 20, 2018Updated 8 years ago
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Simple Example A3C Reinforcement Learning Algorithm in Tensorflow☆13May 23, 2017Updated 9 years ago
- snake-gym is implementation of the classic game snake that is made as an OpenAI gym environment☆24Jul 25, 2024Updated 2 years ago
- A deep reinforcement learning multi-agent algorithm, where a team learns to complete a task and communicate between agents.☆16Jun 1, 2021Updated 5 years ago
- Code for training policies based on paper Coordinated Multi-Agent Imitation Learning☆26Aug 7, 2017Updated 9 years ago
- An PyTorch implementation of "Importance Weighted Actor-Learner Architectures" https://arxiv.org/abs/1802.01561☆13Jan 6, 2021Updated 5 years ago
- Code accompanying paper "Coordinated Proximal Policy Optimization"☆10Mar 26, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- MSc Informatics dissertation project - University of Edinburgh: Curiosity in Multi-Agent Reinforcement Learning☆13Aug 16, 2019Updated 7 years ago
- ☆12Nov 28, 2015Updated 10 years ago
- Multi-Modal Imitation Learning in Partially Observable Environments☆14Sep 5, 2020Updated 6 years ago
- Public implementation of "Encoding Human Domain Knowledge to Warm Start Reinforcement Learning" from AAAI'21☆19Mar 5, 2024Updated 2 years ago
- ☆12Jun 17, 2022Updated 4 years ago
- ☆16Mar 24, 2023Updated 3 years ago
- a collection of DRL-repo in Github☆15Oct 21, 2020Updated 5 years ago
- We reproduced DeepMind's results and implement a meta-learning (MLSH) agent which can generalize across minigames.☆29Mar 30, 2021Updated 5 years ago
- Transplant a implementation of MADDPG to the environment provided by openAI (multiagent-particle-envs).☆21Mar 19, 2018Updated 8 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- A distributed GPU-centric experience replay system for large AI models.☆20Aug 1, 2023Updated 3 years ago
- Hands-on tutorial about Meta RL and GP-MPC at the RL4AA'24 workshop.☆15Apr 20, 2026Updated 5 months ago
- ☆16Dec 15, 2021Updated 4 years ago
- Efficient Adversarial Training without Attacking: Worst-Case-Aware Robust Reinforcement Learning☆30Sep 13, 2023Updated 3 years ago
- A Test-Implementation of the IMPALA algorithm (by deepmind 2018)☆36Mar 16, 2018Updated 8 years ago
- ☆13Aug 15, 2020Updated 6 years ago
- Source code for our NIPS 2017 paper, InfoGAIL: Interpretable Imitation Learning from Visual Demonstrations☆41Nov 16, 2017Updated 8 years ago
- Reinforcement Learning and Transfer Learning based StarCraft Micromanagement☆46Nov 16, 2017Updated 8 years ago
- Learning Individual Intrinsic Reward in MARL☆65Dec 8, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is a Python realization for Milan Korda and Igor Mezic's paper Linear predictors for nonlinear dynamical systems: Koopman operator m…☆14Apr 5, 2023Updated 3 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- Generative adversarial imitation learning on NGSIM I-80 Dataset.☆18Mar 14, 2023Updated 3 years ago
- This repo is related to Deep Policy search using MPC.☆20Jun 1, 2022Updated 4 years ago
- Traffic Signal Control Competition☆41Apr 17, 2019Updated 7 years ago
- My solution code to parallel architecture and programming Spring 2016☆12Aug 15, 2016Updated 10 years ago
- Robust structure identification and room segmentation of cluttered indoor environments from occupancy grid maps☆24Jun 27, 2023Updated 3 years ago