Intrinsic Motivation and Automatic Curricula via Asymmetric Self-Play
☆14May 1, 2018Updated 8 years ago
Alternatives and similar repositories for selfplay
Users that are interested in selfplay are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of Memory Augmented Self-Play☆52Oct 26, 2020Updated 5 years ago
- Pixyz Tutorial in RL Architecture Study Group☆11Apr 25, 2019Updated 7 years ago
- ☆33Oct 17, 2018Updated 7 years ago
- The code accompaniment for the CoRL 2020 paper: A User's Guide to Calibrating Robotics Simulators (https://arxiv.org/abs/2011.08985), fro…☆30Nov 20, 2020Updated 5 years ago
- A minimal toolset for running UI applications within docker isolated X11 environment☆16Jan 11, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- ☆16Aug 12, 2022Updated 4 years ago
- ☆19Feb 18, 2026Updated 6 months ago
- 🔍 Codebase for the ICML '20 paper "Ready Policy One: World Building Through Active Learning" (arxiv: 2002.02693)☆18Jul 6, 2023Updated 3 years ago
- ☆33Jun 14, 2018Updated 8 years ago
- AI LaBuddyは、大規模言語モデルや深層学習を駆使して、研究室のさまざまなタスクを自動化・効率化するためのツール集です。研究者が行う実験や分析、論文の執筆や資料作成を支援し、Slack上でのコミュニケーションを活性化することができます☆22Jul 2, 2023Updated 3 years ago
- Open source demo for the paper Learning to Score Behaviors for Guided Policy Optimization☆24Jun 24, 2020Updated 6 years ago
- ☆19Oct 13, 2021Updated 4 years ago
- 🍞 Bake your robot data into ML-ready formats☆19Feb 25, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- using information theory to encourage agents to cooperate and compete☆19Oct 4, 2018Updated 7 years ago
- PGQ is an approach to combine Policy Gradient and Q-Learning. This repository will contain an implementation of PGQ.☆15Mar 9, 2017Updated 9 years ago
- ☆28Oct 9, 2017Updated 8 years ago
- Code for Environment Probing Interaction Policies [ICLR 2019]☆30Jun 17, 2019Updated 7 years ago
- Latent World Models For Intrinsically Motivated Exploration | Official repository☆23Apr 28, 2021Updated 5 years ago
- Learning from Trajectories via Subgoal Discovery☆12Dec 10, 2020Updated 5 years ago
- ☆10Dec 9, 2021Updated 4 years ago
- Codebase for Efficient yet simple Reinforcement Learning Research Framework☆28Jan 14, 2023Updated 3 years ago
- Code for ICRA2021 "Policy Transfer via Kinematic Domain Randomization and Adaptation"☆12Apr 28, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- MADDPG agent with collaboration and competition☆12Nov 9, 2018Updated 7 years ago
- Surprise-based intrinsic motivation for deep reinforcement learning☆21Mar 6, 2017Updated 9 years ago
- Deep PILCO PyTorch Implementation☆15Mar 25, 2023Updated 3 years ago
- We investigate the effect of populations on finding good solutions to the robust MDP☆29Mar 27, 2021Updated 5 years ago
- MVE: model-based value estimation☆11Jul 30, 2018Updated 8 years ago
- ☆16Jul 31, 2025Updated last year
- ☆38Mar 10, 2022Updated 4 years ago
- A standalone library to randomize various OpenAI Gym Environments☆66Sep 29, 2019Updated 6 years ago
- ICLR 2020 Meta Reinforcement Learning with Autonomous Inference of Subtask Dependencies☆18Jul 16, 2020Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆43Jul 31, 2026Updated 3 weeks ago
- ☆16Oct 3, 2022Updated 3 years ago
- Unofficial Re-implementation of "Dream to Control: Learning Behaviors by Latent Imagination" (https://arxiv.org/abs/1912.01603 ) with PyT…☆32Aug 22, 2020Updated 6 years ago
- A deep reinforcement learning multi-agent algorithm, where a team learns to complete a task and communicate between agents.☆16Jun 1, 2021Updated 5 years ago
- ☆43Oct 9, 2020Updated 5 years ago
- DeepNC: Deep Generative Network Completion☆10Dec 1, 2020Updated 5 years ago
- A System for Morphology-Task Generalization via Unified Representation and Behavior Distillation (ICLR2023)☆14Feb 3, 2023Updated 3 years ago