Experimenting with meta-learning approaches to opponent modelling in MARL. Building upon previous public implementations of MADDPG and M3DDPG.
☆14Apr 26, 2022Updated 4 years ago
Alternatives and similar repositories for LeMOL
Users that are interested in LeMOL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14May 10, 2021Updated 5 years ago
- Official pytorch implementation of the paper <Model-based Multi-agent Policy Optimization with Adaptive Opponent-wise Rollouts>.☆24Nov 22, 2025Updated 10 months ago
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- 论文UAV relay in VANETs against smart jamming with reinforcement learning DQN版本☆25Oct 31, 2021Updated 4 years ago
- On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning☆16Apr 30, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Remote sensing Image Captioning is a special case of Image Captioning which solves the difficulties in processing the remote sensing imag…☆12Jun 16, 2021Updated 5 years ago
- Transplant a implementation of MADDPG to the environment provided by openAI (multiagent-particle-envs).☆21Mar 19, 2018Updated 8 years ago
- ☆13Oct 11, 2022Updated 3 years ago
- This is the code for Q-value Path Decomposition for Deep Multiagent Reinforcement Learning (NeurIPS 2019).☆12May 20, 2019Updated 7 years ago
- Appendix and Code for Modelling Bounded Rationality in Multi-Agent Interactions by Generalized Recursive Reasoning☆14Dec 8, 2022Updated 3 years ago
- Official Implementation for Quality-Similar Diversity via Population Based Reinforcement Learning☆19Dec 26, 2025Updated 9 months ago
- Approximate Message Passing based on Donoho's original paper☆13Jun 28, 2021Updated 5 years ago
- Fundamental of AI course which focuses on search, multiagents, mdp and reinforcement learning algorithms.☆13Oct 29, 2022Updated 3 years ago
- Codebase for BRDiv: Diverse teammate generation for ad hoc teamwork☆13May 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆10Apr 7, 2021Updated 5 years ago
- ☆18Nov 23, 2022Updated 3 years ago
- [ICLR 2023] The official code for paper "Guarded Policy Optimization with Imperfect Online Demonstrations"☆14Apr 30, 2023Updated 3 years ago
- Source code for "A Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning" (ICML 2021)☆34Oct 6, 2022Updated 3 years ago
- Code for Towards Unifying Behavioral and Response Diversity for Open-ended Learning in Zero-sum Games☆24Feb 27, 2022Updated 4 years ago
- Repo for the Greedy when Sure and Conservative when Uncertain about the Opponents (GSCU)☆25Aug 4, 2022Updated 4 years ago
- A PyTorch implementation of ResNet-preact☆12Aug 5, 2019Updated 7 years ago
- ☆14May 26, 2021Updated 5 years ago
- 📈 如何用深度强化学习自动炒股☆13Mar 31, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Soccer toy example simulator used in Reinforcement Learning☆12Mar 11, 2018Updated 8 years ago
- [ICML 2021] DFAC Framework: Factorizing the Value Function via Quantile Mixture for Multi-Agent Distributional Q-Learning☆31Jun 1, 2023Updated 3 years ago
- A Dataset For RSSI Analysis☆17Oct 29, 2021Updated 4 years ago
- Undergraduate Thesis.☆10Apr 13, 2025Updated last year
- A UAVs empowered MEC SYSTEM☆10Sep 22, 2023Updated 3 years ago
- Codes for the paper "SIDE: State Inference for Partially Observable Cooperative Multi-Agent Reinforcement Learning"☆11Jun 24, 2022Updated 4 years ago
- Taylor moment expansion in Python (JaX and SymPy) and Matlab☆11Nov 26, 2024Updated last year
- Advanced_Data_Integration_Project☆11Jul 31, 2018Updated 8 years ago
- Implementation of vanilla stochaistic (categorical) policy gradient algorithm to play cartpole.☆16Apr 1, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Pytorch Implementation of Deep Kalman Filter☆12Sep 30, 2025Updated 11 months ago
- Software Repository accompanying the paper "Ensemble Kalman Filter optimizing Deep NeuralNetworks: An alternative approach to non-perform…☆11Jan 24, 2022Updated 4 years ago
- Code to accompany our International Conference on Robotics and Automation (ICRA) paper entitled - Using variable natural environment brai…☆15Jun 20, 2024Updated 2 years ago
- Random parameter environments using gym 0.7.4 and mujoco-py 0.5.7☆20Feb 14, 2019Updated 7 years ago
- Code for the paper "Symmetric Machine Theory of Mind", presented at ICML 2022.☆12Jul 18, 2022Updated 4 years ago
- This repo contains the ToMnet+ model for preference inference. Developed by Yun-Shiuan, Edwinn, Hsin-Yi, and Elaine.☆11Feb 24, 2023Updated 3 years ago
- Source code for ICML 2023 paper "Competing for Shareable Arms in Multi-Player Multi-Armed Bandits"☆10May 14, 2024Updated 2 years ago