Code for Model-Free Opponent Shaping (ICML 2022)
☆24Nov 18, 2022Updated 3 years ago
Alternatives and similar repositories for model-free-opponent-shaping
Users that are interested in model-free-opponent-shaping are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code release for Learning with Opponent-Learning Awareness and variations.☆152Apr 13, 2023Updated 3 years ago
- ☆15Updated this week
- PyTorch Implementation of the Sequential Multiagent Rollout algorithm☆11Jun 28, 2024Updated 2 years ago
- MARS is shortened for Multi-Agent Research Studio, a library for mulit-agent reinforcement learning research.☆52Mar 8, 2024Updated 2 years ago
- Repo for the Greedy when Sure and Conservative when Uncertain about the Opponents (GSCU)☆25Aug 4, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Efficient baselines for autocurricula in JAX.☆214Aug 24, 2024Updated last year
- Source code of "Variational Imitation Learning with Diverse-quality Demonstrations" in ICML 2020. This github repository includes python …☆20Aug 16, 2021Updated 5 years ago
- Code and data for the paper "Bridging RL Theory and Practice with the Effective Horizon"☆50Jun 26, 2024Updated 2 years ago
- Pytorch implementation of LOLA (https://arxiv.org/abs/1709.04326) using DiCE (https://arxiv.org/abs/1802.05098)☆98Aug 21, 2018Updated 7 years ago
- ☆47May 21, 2024Updated 2 years ago
- Cost-aware Bayesian optimization via the Pandora's box Gittins index☆13Aug 8, 2025Updated last year
- Implementation for ICML 16 paper "Deep reinforcement learning with opponent modeling"☆71Apr 15, 2026Updated 4 months ago
- Code for Discovered Policy Optimisation (NeurIPS 2022)☆12Jun 15, 2023Updated 3 years ago
- Bayesian Optimization Meets Bayesian Optimal Stopping☆32Oct 24, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Learning to Do or Learning While Doing: Reinforcement Learning and Bayesian Optimisation for Online Continuous Tuning…☆14Nov 15, 2023Updated 2 years ago
- Code for magnetic mirror descent.☆20Oct 5, 2023Updated 2 years ago
- Reinforcement learning on general 2D physics environments in JAX. ICLR 2025 Oral.☆266May 21, 2026Updated 2 months ago
- Implementations of Temporal Difference InfoNCE (TD InfoNCE)☆35Nov 13, 2023Updated 2 years ago
- Implementation of "Active Exploration for Inverse Reinforcement Learning (AceIRL), NeurIPS 2022.☆14Oct 12, 2022Updated 3 years ago
- ☆10Apr 13, 2023Updated 3 years ago
- ☆15Sep 22, 2023Updated 2 years ago
- Neural MMO - A Massively Multiagent Environment for Artificial Intelligence Research☆15May 30, 2024Updated 2 years ago
- JAX/Haiku implementation of "Auction Learning as a Two-Player Game"☆11Jul 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"☆35Dec 14, 2023Updated 2 years ago
- Asymmetric methods for partially observable reinforcement learning☆10Jun 9, 2025Updated last year
- 论文Reinforcement Learning of Sequential Price Mechanisms的复现☆12Nov 3, 2022Updated 3 years ago
- Predictive Coding for Locally-Linear Control (ICML-2020)☆18Jul 22, 2024Updated 2 years ago
- Official Repository for "Agent Modelling under Partial Observability for Deep Reinforcement Learning"☆43Oct 5, 2022Updated 3 years ago
- Drop-in environment replacements that make your RL algorithm train faster.☆22Jun 19, 2024Updated 2 years ago
- A tool for aggregating and plotting MARL experiment data.☆86Apr 13, 2026Updated 4 months ago
- Your favourite classical machine learning algos on the GPU/TPU☆23Dec 14, 2025Updated 8 months ago
- Advanced_Data_Integration_Project☆11Jul 31, 2018Updated 8 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code for GLAT (Global Local Transformer), ECCV 2020 "Learning Visual Commonsense for Robust Scene Graph Generation"☆11Dec 16, 2020Updated 5 years ago
- Robust Reinforcement Learning Benchmark☆13Sep 22, 2024Updated last year
- ☆29Mar 16, 2023Updated 3 years ago
- The code used to power DeepRole☆38Nov 21, 2022Updated 3 years ago
- ☆13Mar 12, 2024Updated 2 years ago
- Advantage Alignment Algorithms (ICLR 2025 oral)☆20Apr 7, 2025Updated last year
- ☆12Apr 25, 2022Updated 4 years ago