DecentralizedLearning
☆26Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for MAMBPO
Users that are interested in MAMBPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Mar 25, 2025Updated last year
- Official pytorch implementation of the paper <Model-based Multi-agent Policy Optimization with Adaptive Opponent-wise Rollouts>.☆23Nov 22, 2025Updated 8 months ago
- Cooperative Multi-goal Multi-stage Multi-agent Reinforcement Learning☆58Jun 13, 2022Updated 4 years ago
- Code repository for "N-agent Ad Hoc Teamwork" paper (Wang et al., Neurips 2024).☆29Oct 2, 2025Updated 9 months ago
- ☆12Jan 30, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Multi-task Multi-agent Soft Actor Critic for SMAC☆15Jan 18, 2022Updated 4 years ago
- A Pytorch Implementation of Multi Agent Soft Actor Critic☆44Jan 29, 2019Updated 7 years ago
- Codes accompanying the paper "ROMA: Multi-Agent Reinforcement Learning with Emergent Roles" (ICML 2020 https://arxiv.org/abs/2003.08039)☆171Dec 8, 2022Updated 3 years ago
- ☆34Dec 8, 2022Updated 3 years ago
- Codes accompanying the paper "Believe What You See: Implicit Constraint Approach for Offline Multi-Agent Reinforcement Learning" (NeurIPS…☆76Oct 18, 2022Updated 3 years ago
- [ICML 2021] DFAC Framework: Factorizing the Value Function via Quantile Mixture for Multi-Agent Distributional Q-Learning☆31Jun 1, 2023Updated 3 years ago
- Lipschitz Lifelong RL☆11Nov 6, 2020Updated 5 years ago
- Implementation of DyMA-CL, MARL algorithm☆30Apr 18, 2020Updated 6 years ago
- The implementation of ICLR 2023 paper "Discovering Generalizable Multi-agent Coordination Skills from Multi-task Offline Data".☆45Oct 31, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Safe Multi-Agent MuJoCo benchmark for safe multi-agent reinforcement learning research.☆77Jun 13, 2024Updated 2 years ago
- Official Repository for "Agent Modelling under Partial Observability for Deep Reinforcement Learning"☆43Oct 5, 2022Updated 3 years ago
- Learning Task Embeddings for Teamwork Adaptation in Multi-Agent Reinforcement Learning☆15Apr 25, 2024Updated 2 years ago
- Even Sparser Graph Transformers☆13Dec 4, 2024Updated last year
- ☆43Feb 27, 2024Updated 2 years ago
- ☆30Aug 20, 2021Updated 4 years ago
- Codebase for the Graph-based Policy Learning algorithm, which is designed for learning policies to solve the open ad hoc teamwork problem…☆33Mar 31, 2021Updated 5 years ago
- ☆12Jul 6, 2024Updated 2 years ago
- Implementation for mSAC methods in PyTorch☆42Oct 10, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Modular-HER is revised from OpenAI baselines and supports many improvements for Hindsight Experience Replay as modules.☆17Jun 23, 2021Updated 5 years ago
- Exploring algorithms in the domain of offline reinforcement learning (REM, Ensemble-DQN, DQN, ...)☆17Jul 7, 2020Updated 6 years ago
- Pytorch implementation of the MARL algorithm, MADDPG, which correspondings to the paper "Multi-Agent Actor-Critic for Mixed Cooperative-C…☆683Jul 16, 2022Updated 4 years ago
- RUDDER: Return Decomposition for Delayed Rewards☆49Sep 17, 2020Updated 5 years ago
- ☆120Apr 15, 2023Updated 3 years ago
- ☆13Nov 22, 2022Updated 3 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- 硕士毕业论文代码 深度强化学习☆10Apr 4, 2020Updated 6 years ago
- Code to reproduce experiments from:☆10Dec 11, 2020Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Codebase for Mechanistic Mode Connectivity☆13Jul 14, 2023Updated 3 years ago
- The GRITSBot's circuit board designs, bill of materials, hardware specifications, and 3D design files☆25Dec 5, 2016Updated 9 years ago
- Implementation of Stein Variational Gradient Descent with TensorFlow 2.0☆12Sep 11, 2019Updated 6 years ago
- Reactive crowd simulator used in the IROS 2020 paper: "L2B: Learning to Balance the Safety-Efficiency Trade-off in Interactive Crowd-awar…☆24Dec 19, 2020Updated 5 years ago
- Hierarchical Cooperative Multi-Agent Reinforcement Learning with Skill Discovery☆113Jun 13, 2022Updated 4 years ago
- Terminal sliding mode control for an autonomous motorcycle (Author: Zhenyu Wan, Yicong Xu, Yun Qin)☆21Sep 27, 2018Updated 7 years ago
- A collection of RL algorithms written in JAX.☆105Jul 5, 2022Updated 4 years ago