[AAAI 2023 Oral] Official code for "PiCor: Multi-Task Deep Reinforcement Learning with Policy Correction".
☆21Jul 26, 2025Updated last year
Alternatives and similar repositories for PiCor
Users that are interested in PiCor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI 2025 Oral] Official code for "RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors"☆34Feb 15, 2025Updated last year
- This repository is the official implementation of ZSC-Eval: An Evaluation Toolkit and Benchmark for Multi-agent Zero-shot Coordination. P…☆56Nov 22, 2025Updated 8 months ago
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- ☆32Apr 12, 2026Updated 4 months ago
- Codebase for [Order Matters: Agent-by-agent Policy Optimization](https://openreview.net/forum?id=Q-neeWNVv1)☆32Nov 22, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Overcooked human-AI experiment platform☆41Dec 21, 2023Updated 2 years ago
- FlowBotHD: History-Aware Diffuser Handling Ambiguities in Articulated Objects Manipulation☆13Dec 13, 2024Updated last year
- Code for "On the Utility of Learning about Humans for Human-AI Coordination"☆112Apr 17, 2023Updated 3 years ago
- A large-scale multi-modal pre-trained model☆134Feb 7, 2023Updated 3 years ago
- ☆12Feb 17, 2022Updated 4 years ago
- An NLP research and data collection platform.☆17Jul 4, 2026Updated last month
- multiagent-gail working with multiagent-particle-env-v2 (which was modified by magail authors)☆13Aug 17, 2019Updated 6 years ago
- An extremely light weight tiny-YOLO inference engine targeted towards OpenCL hardware.☆16Oct 15, 2017Updated 8 years ago
- This is the official implementation of paper "Leveraging Dual Process Theory in Language Agent Framework for Simultaneous Human-AI Collab…☆61Nov 22, 2025Updated 8 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Code for "When Should We Prefer State-to-Visual DAgger Over Visual Reinforcement Learning?"☆14Dec 19, 2024Updated last year
- ☆15Jun 28, 2022Updated 4 years ago
- ☆13Oct 11, 2022Updated 3 years ago
- Overcooked-AI Experiment Psiturk Demo (for MTurk experiments)☆13May 10, 2021Updated 5 years ago
- ☆19Apr 8, 2025Updated last year
- [SIGKDD 2023] HardSATGEN: Understanding the Difficulty of Hard SAT Formula Generation and A Strong Structure-Hardness-Aware Baseline☆23Jun 16, 2023Updated 3 years ago
- Official implementation for "ST-GREED: Space-Time Generalized EntropicDifferences for Frame Rate Dependent VideoQuality"☆17Jun 22, 2022Updated 4 years ago
- [SIGIR 2024] TRAD: Enhancing LLM Agents with Step-Wise Thought Retrieval and Aligned Decision☆20Mar 28, 2024Updated 2 years ago
- Official implementation for the paper, StackEval: Benchmarking LLMs in Coding Assistance, https://arxiv.org/abs/2412.05288☆21Oct 30, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A distributed GPU-centric experience replay system for large AI models.☆19Aug 1, 2023Updated 3 years ago
- The implementation of the paper "Where2Explore: Few-shot Affordance Learning for Unseen Novel Categories of Articulated Objects". [NeurIP…☆15Jun 13, 2025Updated last year
- RL environment replicating the werewolf game to study emergent communication☆20May 25, 2023Updated 3 years ago
- ReDMan is an open-source simulation platform that provides a standardized implementation of safe RL algorithms for Reliable Dexterous Man…☆30May 2, 2023Updated 3 years ago
- The official repository of "MemMA: Coordinating the Memory Cycle through Multi-Agent Reasoning and In-Situ Self-Evolution".☆19Mar 20, 2026Updated 4 months ago
- Telegram bot which sends alerts when new papers, articles, books, etc. related to your keywords are released on Google Scholar or arXiv☆12Feb 26, 2022Updated 4 years ago
- This is an implementation of DeepStack for No Limit Texas Hold'em, extended from DeepStack-Leduc.☆26Jun 16, 2019Updated 7 years ago
- This is a repository for Hidden-utility Self-Play.☆27Jul 27, 2023Updated 3 years ago
- ☆18Aug 3, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- test code pcc in 3dg☆19Nov 20, 2017Updated 8 years ago
- Implemention of the Decision-Pretrained Transformer (DPT) from the paper Supervised Pretraining Can Learn In-Context Reinforcement Learni…☆80May 28, 2024Updated 2 years ago
- Repo for the Greedy when Sure and Conservative when Uncertain about the Opponents (GSCU)☆25Aug 4, 2022Updated 4 years ago
- CookingZoo: a gym-cooking derivative to simulate a complex cooking environment☆22Dec 6, 2024Updated last year
- [ICLR 2022 Spotlight] Multi-Stage Episodic Control for Strategic Exploration in Text Games☆15Feb 8, 2026Updated 6 months ago
- ☆13Sep 14, 2021Updated 4 years ago
- ☆15Mar 26, 2024Updated 2 years ago