This is a repository for Hidden-utility Self-Play.
☆27Jul 27, 2023Updated 3 years ago
Alternatives and similar repositories for HSP
Users that are interested in HSP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15May 4, 2024Updated 2 years ago
- Maximum Entropy Population Based Training for Zero-Shot Human-AI Coordination☆26Nov 29, 2022Updated 3 years ago
- Overcooked human-AI experiment platform☆41Dec 21, 2023Updated 2 years ago
- This repository is the official implementation of ZSC-Eval: An Evaluation Toolkit and Benchmark for Multi-agent Zero-shot Coordination. P…☆58Nov 22, 2025Updated 9 months ago
- Offline Multi-Agent Reinforcement Learning Implementations: Solving Overcooked Game with Data-Driven Method☆48Sep 11, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆12Jan 4, 2024Updated 2 years ago
- a jax benchmark for ad hoc teamwork☆24Updated this week
- PantheonRL is a package for training and testing multi-agent reinforcement learning environments. PantheonRL supports cross-play, fine-tu…☆160Nov 6, 2023Updated 2 years ago
- ☆14Jul 12, 2021Updated 5 years ago
- The implementation of ICLR 2023 paper "Discovering Generalizable Multi-agent Coordination Skills from Multi-task Offline Data".☆45Oct 31, 2024Updated last year
- Collection of RL Environments built using Madrona☆40Aug 11, 2023Updated 3 years ago
- ☆16Jul 16, 2024Updated 2 years ago
- This repository contains reference implementation for multi-LLM ToM paper (accepted to EMNLP 2023), Theory of Mind for Multi-Agent Collab…☆20Jun 11, 2024Updated 2 years ago
- Code for "On the Utility of Learning about Humans for Human-AI Coordination"☆112Apr 17, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This is a Human-like Upper-limb Motion Planner (HUMP) for the generation of arm-hand movements in humanoid robots.☆12Mar 4, 2022Updated 4 years ago
- Official implementation of the paper "On the Importance of Environments in Human-Robot Coordination", published in RSS 2021.☆16May 1, 2024Updated 2 years ago
- An environment for table-carrying, a joint-action cooperative task.☆10Jan 8, 2024Updated 2 years ago
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- [ICML 2021] DFAC Framework: Factorizing the Value Function via Quantile Mixture for Multi-Agent Distributional Q-Learning☆31Jun 1, 2023Updated 3 years ago
- ☆16Apr 6, 2022Updated 4 years ago
- Offline Policy Evaluation via Adaptive Weighting with Data from Contextual Bandits☆11Oct 21, 2024Updated last year
- Unveiling the Knowledge of Hindu Scriptures☆13Mar 31, 2025Updated last year
- An NLP research and data collection platform.☆17Jul 4, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- This is a project for creating and using IL datasets based on HuggingFace weights with multithreads for performance, and benchmarking☆13Aug 5, 2026Updated last month
- Pytorch code for "Learning Guidance Rewards with Trajectory-space Smoothing" (NeurIPS 2020)☆12Jul 7, 2021Updated 5 years ago
- Repository with environment and training scripts for paper "Cross-Environment-Cooperation Enables Zero-shot Multi-agent Cooperation"☆22Sep 12, 2025Updated last year
- A benchmark environment for fully cooperative human-AI performance.☆1,007Mar 22, 2025Updated last year
- A beginner-friendly repository on Deep Reinforcement Learning (RL), written in PyTorch.☆27Mar 19, 2026Updated 6 months ago
- Installing Julia and Python environments☆14May 29, 2024Updated 2 years ago
- Companion code for ICML 2022 paper "Imitation Learning by Estimating Expertise of Demonstrators"☆12Jul 5, 2023Updated 3 years ago
- First EEG course in WS2020 in Stuttgart☆15Jun 24, 2021Updated 5 years ago
- Official Code For EMNLP2025 Findings: {DLPO : Towards a Robust, Efficient, and Generalizable Prompt Optimization Framework from a Deep-Le…☆10Dec 25, 2025Updated 8 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official codebase for Generating Diverse Cooperative Agents by Learning Incompatible Policies (notable-top-25% @ ICLR 2023)☆19May 10, 2024Updated 2 years ago
- League of Legends, Teamfight Tactics Environment for RL(not complete)☆13Jun 20, 2020Updated 6 years ago
- ☆517Dec 28, 2023Updated 2 years ago
- suPER is a collaborative multi-agent RL algorithm☆14Jun 11, 2024Updated 2 years ago
- A list of academic papers that utilize the ManiSkill framework, categorized by year of publication.☆25May 30, 2025Updated last year
- Code for "Traffic Signal Cycle Control with Centralized Critic and Decentralized Actors under Varying Intervention Frequencies"☆14Jun 27, 2025Updated last year
- A GPU-accelerated toolbox for hyperbolic PDEs in a weaker (viscosity) sense. It leverages the integral to the solution of the conservatio…☆16Updated this week