A new paper list for multi-agent reinforcement learning (actively maintained)
☆24Mar 27, 2020Updated 6 years ago
Alternatives and similar repositories for Paper-List-of-MARL
Users that are interested in Paper-List-of-MARL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Mar 16, 2023Updated 3 years ago
- Meta-Reinforcement Learning with Policy Residual Representation☆11Aug 15, 2019Updated 7 years ago
- Personal Repo to keep track of RL papers☆31May 3, 2021Updated 5 years ago
- Simulink Reference example for modeling smart trucks with the intelligence to form a platoon based on certain criteria.☆27Nov 26, 2019Updated 6 years ago
- Benchmark result of different RL algorithms on MetaDrive environments, including Multi-agent RL (IPPO, centralized critics, CoPO).☆16Oct 25, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆109Feb 10, 2021Updated 5 years ago
- rllab is a framework for developing and evaluating reinforcement learning algorithms, fully compatible with OpenAI Gym.☆14Apr 3, 2017Updated 9 years ago
- Fine-tuned MARL algorithms on SMAC (100% win rates on most scenarios)☆19Aug 20, 2023Updated 3 years ago
- In this paper, we propose Filter Gradient Decent (FGD), an efficient stochastic optimization algorithm that makes a consistent estimation…☆12May 18, 2021Updated 5 years ago
- The Knowledge Graph is tool for organizing and sharing your knowledge online. Use it to keep track of what you have learned and share you…☆14Feb 28, 2023Updated 3 years ago
- This ROS node includes C++ implementations for extracting OpenStreetMaps(OSM), performing planning using latitude/longitude or 2D relativ…☆13Feb 22, 2025Updated last year
- SIR, SEIR, and beyond☆10Jul 6, 2023Updated 3 years ago
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- Code for paper 'Learning transferable cooperative behaviors in multi-agent teams' (ICML 2019)☆124Dec 8, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for ICLR 2019 paper: Learning when to Communicate at Scale in Multiagent Cooperative and Competitive Tasks☆232Oct 3, 2023Updated 2 years ago
- A simple program scheduler for your code on different devices.☆12Mar 8, 2026Updated 6 months ago
- Learning Individual Intrinsic Reward in MARL☆65Dec 8, 2022Updated 3 years ago
- [Findings of ACL 2023] Communication Efficient Federated Learning for Multilingual Machine Translation with Adapter☆12Sep 4, 2023Updated 3 years ago
- Code for the paper☆12May 24, 2024Updated 2 years ago
- OmniByteFormer is a generalized Transformer model that can process any type of data by converting it into byte sequences, bypassing tradi…☆17Updated this week
- Multi-objective reinforcement learning for covid-19 control☆12Aug 12, 2021Updated 5 years ago
- Code for NeurIPS 2019 paper "Screening Sinkhorn Algorithm for Regularized Optimal Transport"☆10Feb 10, 2020Updated 6 years ago
- Submission for MAVEN: Multi-Agent Variational Exploration☆58Apr 6, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- OPE Tools based on Empirical Study of Off Policy Policy Estimation paper.☆61Aug 9, 2022Updated 4 years ago
- Code for the paper "Training Binary Neural Networks with Bayesian Learning Rule☆41Jan 4, 2022Updated 4 years ago
- ☆10Nov 4, 2019Updated 6 years ago
- ☆15Jun 22, 2020Updated 6 years ago
- A pathway and collection of resources to learning Jax from beginning to advance.☆11Jan 2, 2021Updated 5 years ago
- We implement MADDPG in a congestion env, and compare with several control groups to highlight the performance of MADDPG☆11Jul 14, 2021Updated 5 years ago
- Official PyTorch implementation of "ACE:Off-Policy Actor-Critic with Causality-Aware Entropy Regularization"☆36May 13, 2024Updated 2 years ago
- pytorch implementation of "Efficient Communication in Multi-Agent Reinforcement Learning via Variance Based Control"☆54Dec 8, 2022Updated 3 years ago
- Code for [NeurIPS'2019 Spotlight] Policy Continuation with Hindsight Inverse Dynamics☆15Jan 7, 2020Updated 6 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆12Aug 15, 2020Updated 6 years ago
- ☆17Nov 27, 2024Updated last year
- Code for SyncTwin: Treatment Effect Estimation with Longitudinal Outcomes (NeurIPS 2021)☆13Nov 30, 2021Updated 4 years ago
- SimPER: A Minimalist Approach to Preference Alignment without Hyperparameters (ICLR 2025)☆17Aug 22, 2025Updated last year
- A public repo for ICML 2021 "Shortest-Path Constrained Reinforcement Learning for Sparse Reward Tasks"☆13Jul 19, 2021Updated 5 years ago
- ☆12Sep 30, 2017Updated 8 years ago
- Threat Network Detection in Online Social Networks☆11Jan 20, 2017Updated 9 years ago