A new paper list for multi-agent reinforcement learning (actively maintained)
☆24Mar 27, 2020Updated 6 years ago
Alternatives and similar repositories for Paper-List-of-MARL
Users that are interested in Paper-List-of-MARL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Mar 16, 2023Updated 3 years ago
- Codes accompanying the paper "Influence-Based Multi-Agent Exploration" (ICLR 2020 spotlight)☆34Mar 16, 2020Updated 6 years ago
- Meta-Reinforcement Learning with Policy Residual Representation☆11Aug 15, 2019Updated 7 years ago
- Simulink Reference example for modeling smart trucks with the intelligence to form a platoon based on certain criteria.☆27Nov 26, 2019Updated 6 years ago
- Benchmark result of different RL algorithms on MetaDrive environments, including Multi-agent RL (IPPO, centralized critics, CoPO).☆16Oct 25, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆109Feb 10, 2021Updated 5 years ago
- Fine-tuned MARL algorithms on SMAC (100% win rates on most scenarios)☆19Aug 20, 2023Updated 3 years ago
- In this paper, we propose Filter Gradient Decent (FGD), an efficient stochastic optimization algorithm that makes a consistent estimation…☆12May 18, 2021Updated 5 years ago
- Source code for Interpretable Reward Redistribution in Reinforcement Learning: A Causal Approach (NeurIPS 2023)☆10Dec 12, 2023Updated 2 years ago
- This ROS node includes C++ implementations for extracting OpenStreetMaps(OSM), performing planning using latitude/longitude or 2D relativ…☆14Feb 22, 2025Updated last year
- SIR, SEIR, and beyond☆10Jul 6, 2023Updated 3 years ago
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- Code for paper 'Learning transferable cooperative behaviors in multi-agent teams' (ICML 2019)☆124Dec 8, 2022Updated 3 years ago
- Automatic generation of architecture-level models for hardware from its RTL design.☆16Apr 12, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for ICLR 2019 paper: Learning when to Communicate at Scale in Multiagent Cooperative and Competitive Tasks☆232Oct 3, 2023Updated 2 years ago
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- 清华大学研究生社会实践系统爬虫☆17Jun 4, 2024Updated 2 years ago
- Learning Individual Intrinsic Reward in MARL☆65Dec 8, 2022Updated 3 years ago
- The Reinforcement-Learning-Related Papers of ICLR 2019☆47May 28, 2019Updated 7 years ago
- OmniByteFormer is a generalized Transformer model that can process any type of data by converting it into byte sequences, bypassing tradi…☆16Aug 24, 2026Updated last week
- Multi-objective reinforcement learning for covid-19 control☆12Aug 12, 2021Updated 5 years ago
- Code for NeurIPS 2019 paper "Screening Sinkhorn Algorithm for Regularized Optimal Transport"☆10Feb 10, 2020Updated 6 years ago
- Submission for MAVEN: Multi-Agent Variational Exploration☆58Apr 6, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- OPE Tools based on Empirical Study of Off Policy Policy Estimation paper.☆61Aug 9, 2022Updated 4 years ago
- Code for the paper "Training Binary Neural Networks with Bayesian Learning Rule☆41Jan 4, 2022Updated 4 years ago
- 多智能体学习库☆22Dec 28, 2021Updated 4 years ago
- ☆15Jun 22, 2020Updated 6 years ago
- [EMNLP 2021] PyTorch Implementation of Contrastive Domain Adaptation for Question Answering using Limited Text Corpora☆14Jul 4, 2023Updated 3 years ago
- We implement MADDPG in a congestion env, and compare with several control groups to highlight the performance of MADDPG☆11Jul 14, 2021Updated 5 years ago
- Official PyTorch implementation of "ACE:Off-Policy Actor-Critic with Causality-Aware Entropy Regularization"☆35May 13, 2024Updated 2 years ago
- pytorch implementation of "Efficient Communication in Multi-Agent Reinforcement Learning via Variance Based Control"☆54Dec 8, 2022Updated 3 years ago
- CARMA Streets is a component of CARMA ecosystem, which enables such a coordination among different transportation users. This component p…☆11Aug 21, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for [NeurIPS'2019 Spotlight] Policy Continuation with Hindsight Inverse Dynamics☆15Jan 7, 2020Updated 6 years ago
- ☆17Nov 27, 2024Updated last year
- A public repo for ICML 2021 "Shortest-Path Constrained Reinforcement Learning for Sparse Reward Tasks"☆13Jul 19, 2021Updated 5 years ago
- ☆12Sep 30, 2017Updated 8 years ago
- Threat Network Detection in Online Social Networks☆11Jan 20, 2017Updated 9 years ago
- Repo containing code for multi-agent deep reinforcement learning (MADRL).☆752Jul 7, 2026Updated last month
- A Frank-Wolfe Framework for Efficient and Effective Adversarial Attacks (AAAI'20)☆11Jun 10, 2020Updated 6 years ago