☆26Apr 12, 2018Updated 8 years ago
Alternatives and similar repositories for qmix
Users that are interested in qmix are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- QMIX implemented in TensorFlow 2☆17Jun 12, 2021Updated 5 years ago
- qmix☆23May 28, 2020Updated 6 years ago
- Improving upon state of the art cooperative deep reinforcement learning in StarCraft II☆13May 16, 2019Updated 7 years ago
- The project to learn the QMIX.☆13Dec 19, 2019Updated 6 years ago
- ☆17Dec 4, 2019Updated 6 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- a collection of DRL-repo in Github☆15Oct 21, 2020Updated 5 years ago
- ☆10Feb 28, 2019Updated 7 years ago
- Code for the papers "Modeling the Second Player in Distributionally Robust Optimization" and "Distributionally Robust Models with Paramet…☆29Apr 14, 2022Updated 4 years ago
- paper on dexpilot☆15Oct 14, 2019Updated 6 years ago
- There will be updates later☆87May 13, 2019Updated 7 years ago
- ☆44Feb 12, 2020Updated 6 years ago
- ☆44Oct 27, 2018Updated 7 years ago
- ☆15Jan 27, 2025Updated last year
- ☆10May 23, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Random parameter environments using gym 0.7.4 and mujoco-py 0.5.7☆20Feb 14, 2019Updated 7 years ago
- Assignments for CS294-112.☆30Sep 11, 2019Updated 6 years ago
- ☆14Feb 4, 2022Updated 4 years ago
- This repository contains implementations of the paper, Bayesian Model-Agnostic Meta-Learning.☆20Jan 19, 2023Updated 3 years ago
- Knowledge Distillation Algorithms implemented with PyTorch☆17Jul 23, 2019Updated 7 years ago
- ☆12Feb 20, 2021Updated 5 years ago
- A Python 3 implementation of a Blockchain with a PBFT conscientious mechanism.☆17May 14, 2026Updated 2 months ago
- 🔬 Absolutely comfort lab for me to work around with my own AIs and to empirically observe how powerful and impactful these technologies …☆29Jan 12, 2025Updated last year
- ☆10Jul 13, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Multi-Modal Imitation Learning in Partially Observable Environments☆14Sep 5, 2020Updated 5 years ago
- Implementation of DeDOL algorithm - Deep Reinforcement Learning based algorithm for Green Security Games with Real Time Information☆16Nov 7, 2019Updated 6 years ago
- Python neighbor-joining library. Goal: Efficient O(n^2) neighbor-joining algorithm.☆12May 5, 2014Updated 12 years ago
- A code implementation for our arXiv paper "Multi-agent Adhoc Team Play using Decompositional Q function"☆132Aug 14, 2023Updated 2 years ago
- Resilient Multi-Agent Reinforcement Learning☆10Nov 4, 2022Updated 3 years ago
- Implementation of (Learning Continuous Control Policies by Stochastic Value Gradients)[https://arxiv.org/abs/1510.09142]☆25Jan 15, 2022Updated 4 years ago
- Exploration by Random Network Distillation☆15Dec 30, 2018Updated 7 years ago
- BILIBILI.☆15Jan 6, 2019Updated 7 years ago
- ☆12Jun 17, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Using RLLib and PycoLab to explore intelligent cooperative behavior in sequential social dilemmas☆54Dec 8, 2022Updated 3 years ago
- ☆15Dec 31, 2020Updated 5 years ago
- Code for NeurIPS 2021 paper "Offline Reinforcement Learning with Reverse Model-based Imagination"☆20Dec 22, 2021Updated 4 years ago
- Simple one-dimensional data classification using LSTM, CNN, FC in Pytorch☆23Nov 2, 2019Updated 6 years ago
- Pytorch implementation of "FeUdal Networks for Hierarchical Reinforcement Learning" for Montezuma's Revenge☆95Jul 27, 2022Updated 3 years ago
- 浙江大学Beamer模板☆16May 19, 2022Updated 4 years ago
- An implementation of Counterfactual Regret Minimization (CFR) via Temporal Difference (TD) learning☆22May 11, 2013Updated 13 years ago