The AI Arena: A framework for distributed multi-agent reinforcement learning
☆14Aug 5, 2022Updated 4 years ago
Alternatives and similar repositories for ai-arena
Users that are interested in ai-arena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for our GECCO 2021 paper : A Coevolutionary Approach to Deep Multi-agent Reinforcement Learning☆15Oct 4, 2021Updated 4 years ago
- Adaptable Agent Populations via a Generative Model of Policies☆12Oct 14, 2021Updated 4 years ago
- Appendix and Code for Modelling Bounded Rationality in Multi-Agent Interactions by Generalized Recursive Reasoning☆14Dec 8, 2022Updated 3 years ago
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- MaxSum is an algorithm about Distributed Constraint Optimization Problems (DCOPs)☆11Jan 15, 2018Updated 8 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- TS_SPMA: The Tabu Search algorithm for simultaneous scheduling problem of machines and AGVs.☆12Apr 30, 2021Updated 5 years ago
- Official code for "Pretraining Representations For Data-Efficient Reinforcement Learning" (NeurIPS 2021)☆56Jul 27, 2021Updated 5 years ago
- 📅 Production-ready scheduler with async, multithreading and multiprocessing support for Python☆22Jul 6, 2024Updated 2 years ago
- code of paper 《Independent Reinforcement Learning for Weakly Cooperative Multiagent Traffic Control Problem》☆17Dec 14, 2020Updated 5 years ago
- Python package for computing partial information decomposition.☆13Mar 15, 2019Updated 7 years ago
- RL projects including implementation of DQN/DDPG/MADDPG/BicNet on StarCraft II multi-agent learning environment SMAC☆46Feb 7, 2020Updated 6 years ago
- The code of the algorithm proposed in the paper "Deep Inverse Reinforcement Learning for Objective Function Identification in Bidding Mod…☆15Aug 13, 2021Updated 5 years ago
- 硕士毕业论文代码 深度强化学习☆10Apr 4, 2020Updated 6 years ago
- ppx_system is a syntax extension to known operating system at compile time☆12May 9, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code to reproduce experiments from:☆10Dec 11, 2020Updated 5 years ago
- PyTorch implementation of Count-Based Exploration with Neural Density Models☆10Mar 22, 2018Updated 8 years ago
- MishformerLens intends to be a drop-in replacement for TransformerLens that AST patches HuggingFace Transformers rather than implementing…☆10Oct 7, 2024Updated last year
- A StarCraft 2 agent for harvesting resources☆13Jun 12, 2018Updated 8 years ago
- OpenAI gym environments for goal-conditioned and language-conditioned reinforcement learning☆14Jan 27, 2026Updated 7 months ago
- Monte Carlo tree search for the travelling salesman problem (MCTS for the TSP)☆13Jun 18, 2022Updated 4 years ago
- Implements the Messenger environment and EMMA model.☆25Jun 14, 2023Updated 3 years ago
- **Sferes2 module** A unifying modular framework for Quality-Diversity algorithms☆22Nov 6, 2020Updated 5 years ago
- A simple and easy to use implementation of the soft actor-critic algorithm.☆15Sep 2, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Path finding, task scheduling for multiple agv robot☆22Dec 9, 2022Updated 3 years ago
- Reference Python implementation of Lyapunov-based online scheduling for energy minimization in multi-AP wireless-powered mobile edge comp…☆10Aug 3, 2026Updated last month
- Pytorch implementation of the Gato paper from Deepmind☆12Feb 8, 2023Updated 3 years ago
- This is a ROS repository to track an underwater target using a Particle Filter range-only method and the SparusII AUV☆11Nov 27, 2024Updated last year
- Implementation of the Lemke-Howson algorithm for finding MNE☆15Nov 2, 2013Updated 12 years ago
- Soft-QMIX: Integrating Maximum Entropy For Monotonic Value Function Factorization☆16Jul 3, 2024Updated 2 years ago
- Code related to the Neural-Swarm (ICRA 2020, Journal) papers☆29Mar 23, 2022Updated 4 years ago
- Combination of Maskable PPO and Recurrent PPO based on the sb3-contrib repository☆12Feb 22, 2023Updated 3 years ago
- all conf for apps☆15Apr 26, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Apr 7, 2021Updated 5 years ago
- Algorithm that combines QMIX with SAC for Multi-Agent Reinforcement Learning.☆59May 20, 2022Updated 4 years ago
- A powerful keybind library and daemon for Linux.☆11Jul 24, 2022Updated 4 years ago
- Provides a jailbreak experience of AWS DeepRacer, giving us more control over the training/simulation process and RL algorithm tuning☆18Feb 17, 2023Updated 3 years ago
- This is an implementation of the paper Cooperative and Distributed Reinforcement Learning of Drones for Field Coverage by Huy Xuan Pham, …☆20Jun 29, 2020Updated 6 years ago
- Implementation of Diversity Is All You Need (DIAYN) on top of Stable Baselines 3.☆13Jul 11, 2022Updated 4 years ago
- Brutaltester compatible referee for coders strike back☆13Jun 1, 2026Updated 3 months ago