[AAAI 2023 Oral] Official code for "PiCor: Multi-Task Deep Reinforcement Learning with Policy Correction".
☆21Jul 26, 2025Updated 11 months ago
Alternatives and similar repositories for PiCor
Users that are interested in PiCor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [AAAI 2025 Oral] Official code for "RAT: Adversarial Attacks on Deep Reinforcement Agents for Targeted Behaviors"☆34Feb 15, 2025Updated last year
- This repository is the official implementation of ZSC-Eval: An Evaluation Toolkit and Benchmark for Multi-agent Zero-shot Coordination. P…☆56Nov 22, 2025Updated 8 months ago
- ☆15Oct 28, 2024Updated last year
- Exploring techniques to generate diverse conventions in multi-agent settings☆16Nov 14, 2023Updated 2 years ago
- CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR☆21Apr 7, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Codebase for [Order Matters: Agent-by-agent Policy Optimization](https://openreview.net/forum?id=Q-neeWNVv1)☆32Nov 22, 2025Updated 8 months ago
- Overcooked human-AI experiment platform☆41Dec 21, 2023Updated 2 years ago
- FlowBotHD: History-Aware Diffuser Handling Ambiguities in Articulated Objects Manipulation☆13Dec 13, 2024Updated last year
- Code for "On the Utility of Learning about Humans for Human-AI Coordination"☆112Apr 17, 2023Updated 3 years ago
- ☆12Feb 17, 2022Updated 4 years ago
- An NLP research and data collection platform.☆17Jul 4, 2026Updated 2 weeks ago
- multiagent-gail working with multiagent-particle-env-v2 (which was modified by magail authors)☆13Aug 17, 2019Updated 6 years ago
- ☆12Apr 12, 2022Updated 4 years ago
- a toolkit of KartRider☆17Oct 8, 2024Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- The homework of robos learning base.☆11May 23, 2023Updated 3 years ago
- An extremely light weight tiny-YOLO inference engine targeted towards OpenCL hardware.☆16Oct 15, 2017Updated 8 years ago
- This is the official implementation of paper "Leveraging Dual Process Theory in Language Agent Framework for Simultaneous Human-AI Collab…☆61Nov 22, 2025Updated 8 months ago
- ☆15Jun 28, 2022Updated 4 years ago
- Overcooked-AI Experiment Psiturk Demo (for MTurk experiments)☆13May 10, 2021Updated 5 years ago
- ☆18Apr 8, 2025Updated last year
- [SIGKDD 2023] HardSATGEN: Understanding the Difficulty of Hard SAT Formula Generation and A Strong Structure-Hardness-Aware Baseline☆23Jun 16, 2023Updated 3 years ago
- Official implementation for "ST-GREED: Space-Time Generalized EntropicDifferences for Frame Rate Dependent VideoQuality"☆17Jun 22, 2022Updated 4 years ago
- [SIGIR 2024] TRAD: Enhancing LLM Agents with Step-Wise Thought Retrieval and Aligned Decision☆20Mar 28, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official implementation for the paper, StackEval: Benchmarking LLMs in Coding Assistance, https://arxiv.org/abs/2412.05288☆20Oct 30, 2024Updated last year
- A flexible Multi-Agent Reinforcement Learning (MARL) environment for Collective Robotic Construction (CRC) systems☆13Mar 22, 2023Updated 3 years ago
- A distributed GPU-centric experience replay system for large AI models.☆19Aug 1, 2023Updated 2 years ago
- The implementation of the paper "Where2Explore: Few-shot Affordance Learning for Unseen Novel Categories of Articulated Objects". [NeurIP…☆15Jun 13, 2025Updated last year
- A PyTorch implementation of PTSA-MCTS from [Accelerating Monte Carlo Tree Search with Probability Tree State Abstraction].☆16Oct 21, 2023Updated 2 years ago
- RL environment replicating the werewolf game to study emergent communication☆20May 25, 2023Updated 3 years ago
- ReDMan is an open-source simulation platform that provides a standardized implementation of safe RL algorithms for Reliable Dexterous Man…☆29May 2, 2023Updated 3 years ago
- The official repository of "MemMA: Coordinating the Memory Cycle through Multi-Agent Reasoning and In-Situ Self-Evolution".☆19Mar 20, 2026Updated 4 months ago
- ☆18May 28, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The first, open access evaluation dataset for methods to identify bias by word choice and labeling☆26Oct 30, 2025Updated 8 months ago
- DQN with freezing target network in tensorflow on pygame FlappyBird☆11Dec 19, 2018Updated 7 years ago
- This is an implementation of DeepStack for No Limit Texas Hold'em, extended from DeepStack-Leduc.☆26Jun 16, 2019Updated 7 years ago
- A list of robotics related papers accepted by ICLR'25☆25Aug 28, 2025Updated 10 months ago
- ☆18Aug 3, 2022Updated 3 years ago
- test code pcc in 3dg☆19Nov 20, 2017Updated 8 years ago
- Implemention of the Decision-Pretrained Transformer (DPT) from the paper Supervised Pretraining Can Learn In-Context Reinforcement Learni…☆79May 28, 2024Updated 2 years ago