Code for Model-Free Opponent Shaping (ICML 2022)
☆24Nov 18, 2022Updated 3 years ago
Alternatives and similar repositories for model-free-opponent-shaping
Users that are interested in model-free-opponent-shaping are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Scalable Opponent Shaping Experiments in JAX☆27Apr 13, 2024Updated 2 years ago
- Source code for "A Policy Gradient Algorithm for Learning to Learn in Multiagent Reinforcement Learning" (ICML 2021)☆34Oct 6, 2022Updated 3 years ago
- Code release for Learning with Opponent-Learning Awareness and variations.☆152Apr 13, 2023Updated 3 years ago
- ☆15Jul 20, 2026Updated last week
- PyTorch Implementation of the Sequential Multiagent Rollout algorithm☆11Jun 28, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for the paper "Stable Gradients for Stable Learning at Scale in Deep Reinforcement Learning". Great performance in many environments…☆39Oct 24, 2025Updated 9 months ago
- MARS is shortened for Multi-Agent Research Studio, a library for mulit-agent reinforcement learning research.☆51Mar 8, 2024Updated 2 years ago
- Official code of Nash-DQN for paper: Nash-DQN algorithm for two-player zero-sum Markov games, details see our paper: A Deep Reinforcement…☆22Aug 26, 2022Updated 3 years ago
- POPGym Library in JAX☆14Apr 15, 2024Updated 2 years ago
- ☆16Jul 16, 2024Updated 2 years ago
- Efficient baselines for autocurricula in JAX.☆214Aug 24, 2024Updated last year
- Source code of "Variational Imitation Learning with Diverse-quality Demonstrations" in ICML 2020. This github repository includes python …☆20Aug 16, 2021Updated 4 years ago
- Code and data for the paper "Bridging RL Theory and Practice with the Effective Horizon"☆50Jun 26, 2024Updated 2 years ago
- Cost-aware Bayesian optimization via the Pandora's box Gittins index☆13Aug 8, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation for ICML 16 paper "Deep reinforcement learning with opponent modeling"☆71Apr 15, 2026Updated 3 months ago
- Code for Discovered Policy Optimisation (NeurIPS 2022)☆12Jun 15, 2023Updated 3 years ago
- Bayesian Optimization Meets Bayesian Optimal Stopping☆32Oct 24, 2020Updated 5 years ago
- Code for the paper "Learning to Do or Learning While Doing: Reinforcement Learning and Bayesian Optimisation for Online Continuous Tuning…☆14Nov 15, 2023Updated 2 years ago
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago
- Reinforcement learning on general 2D physics environments in JAX. ICLR 2025 Oral.☆263May 21, 2026Updated 2 months ago
- Baselines for gymnax 🤖☆78Apr 3, 2023Updated 3 years ago
- a jax benchmark for ad hoc teamwork☆22Updated this week
- Implementations of Temporal Difference InfoNCE (TD InfoNCE)☆35Nov 13, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation of "Active Exploration for Inverse Reinforcement Learning (AceIRL), NeurIPS 2022.☆14Oct 12, 2022Updated 3 years ago
- An Open-Ended Agentic Simulator☆61Aug 11, 2024Updated last year
- ☆10Apr 13, 2023Updated 3 years ago
- ☆15Sep 22, 2023Updated 2 years ago
- JAX/Haiku implementation of "Auction Learning as a Two-Player Game"☆11Jul 6, 2024Updated 2 years ago
- Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"☆35Dec 14, 2023Updated 2 years ago
- Asymmetric methods for partially observable reinforcement learning☆10Jun 9, 2025Updated last year
- Highly scalable 2D JAX physics engine.☆68Apr 20, 2026Updated 3 months ago
- 论文Reinforcement Learning of Sequential Price Mechanisms的复现☆12Nov 3, 2022Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Predictive Coding for Locally-Linear Control (ICML-2020)☆18Jul 22, 2024Updated 2 years ago
- Drop-in environment replacements that make your RL algorithm train faster.☆22Jun 19, 2024Updated 2 years ago
- A tool for aggregating and plotting MARL experiment data.☆86Apr 13, 2026Updated 3 months ago
- Your favourite classical machine learning algos on the GPU/TPU☆23Dec 14, 2025Updated 7 months ago
- Risk-sensitive Inverse Reinforcement Learning☆11Sep 11, 2019Updated 6 years ago
- Official Code for "Monte Carlo Tree Diffusion for System 2 Planning" and "Fast Monte Carlo Tree Diffusion: 100x Speedup via Parallel and …☆29Aug 21, 2025Updated 11 months ago
- Advanced_Data_Integration_Project☆11Jul 31, 2018Updated 7 years ago