PyTorch implementation of Sample Efficient Actor-Critic with Experience Replay(ACER)
☆16Oct 7, 2020Updated 5 years ago
Alternatives and similar repositories for pytorch-acer
Users that are interested in pytorch-acer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Adaptive Hypernetworks for Multi-Agent RL. NeurIPS 2025.☆25Apr 14, 2026Updated 5 months ago
- Code for the paper Alpha Zero in Continuous Action Space (A0C) (https://arxiv.org/pdf/1805.09613.pdf)☆15Jan 19, 2021Updated 5 years ago
- Proximal Policy Optimization with Stein Control Variates:☆34Feb 12, 2018Updated 8 years ago
- Codes accompanying the paper "DOP: Off-Policy Multi-Agent Decomposed Policy Gradients" (ICLR 2021, https://arxiv.org/abs/2007.12322)☆51Dec 8, 2022Updated 3 years ago
- Contains an implementation of "Imitation Learning via Kernel Mean Embedding (2018, AAAI)"☆11Oct 2, 2018Updated 7 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [NeurIPS'20] Code for the paper "Offline Imitation Learning with a Misspecified Simulator"☆12Nov 24, 2021Updated 4 years ago
- Giving Up Control: Neurons as Reinforcement Learning Agents☆13May 6, 2024Updated 2 years ago
- Actor-critic with experience replay☆257Oct 9, 2022Updated 3 years ago
- Deep Reinforcement Learning for Multi Agent Soccer☆16Dec 15, 2016Updated 9 years ago
- PyTorch implementation of "Sample-efficient Imitation Learning via Generative Adversarial Nets"☆10Nov 22, 2019Updated 6 years ago
- Code of Truly Batch Model-Free Inverse Reinforcement Learning about Multiple Intentions☆13May 22, 2023Updated 3 years ago
- ☆38Dec 26, 2022Updated 3 years ago
- Author's PyTorch implementation of paper "Provably Good Batch Reinforcement Learning Without Great Exploration"☆11Oct 22, 2020Updated 5 years ago
- Efficiently send large arrays across machines☆15Jul 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- LRP for LSTMs, GRUs, and BERT.☆17Jan 16, 2021Updated 5 years ago
- ☆10Mar 13, 2023Updated 3 years ago
- ☆10May 13, 2025Updated last year
- Implementation of "Sample-Efficient Deep Reinforcement Learning via Episodic Backward Update", NeurIPS 2019.☆16Sep 24, 2019Updated 6 years ago
- Source code to the AAAI21 publication Augmenting Policy Learning with Routines Discovered from a Single Demonstration☆17Jan 7, 2021Updated 5 years ago
- ☆24Jan 12, 2021Updated 5 years ago
- implementing Weight Agnostic Neural Networks to Spiking Neural Networks☆10Jan 26, 2021Updated 5 years ago
- Humanoid behavior imitation using Generative Adversarial Imitation Learning (GAIL)☆16Jul 1, 2020Updated 6 years ago
- Exploring algorithms in the domain of offline reinforcement learning (REM, Ensemble-DQN, DQN, ...)☆17Jul 7, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆20Sep 14, 2019Updated 7 years ago
- ☆27Oct 10, 2020Updated 5 years ago
- Distributed Priortized Experience Replay☆10Aug 8, 2018Updated 8 years ago
- The implement of GAIL with pytorch☆14Mar 11, 2020Updated 6 years ago
- Various reinforcement learning algorithms written in Jax + Flax☆26Jun 24, 2023Updated 3 years ago
- GreenAug: Green Screen Augmentation Enables Scene Generalisation in Robotic Manipulation☆14Sep 10, 2024Updated 2 years ago
- The codebase and datasets for the IJCAI 2021 paper "The Surprising Power of Graph Neural Networks with Random Node Initialization".☆22Jun 3, 2021Updated 5 years ago
- Domain-Robust Visual Imitation Learning with Mutual Information Constraints code☆19Mar 1, 2021Updated 5 years ago
- Deep Reinforcement Learning by using Phasic Policy Gradient in Pytorch & Tensorflow☆20Oct 5, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- "Detecting Extrapolation with Local Ensembles" by David Madras, James Atwood, and Alex D'Amour☆13Sep 25, 2020Updated 5 years ago
- Code for "SMIX(λ): Enhancing Centralized Value Functions for Cooperative Multi-Agent Reinforcement Learning" AAAI 2020☆27Dec 8, 2022Updated 3 years ago
- Tensorflow Implementation of adversarial learning based adversarial example generator☆10Jan 31, 2018Updated 8 years ago
- ☆13Jan 31, 2026Updated 7 months ago
- Wasserstein Distance guided Adversarial Imitation Learning (WDAIL) with Reward Shape Exploration☆19Feb 9, 2021Updated 5 years ago
- Pytorch code for "State-only Imitation with Transition Dynamics Mismatch" (ICLR 2020)☆20Feb 29, 2020Updated 6 years ago
- Code for Paper "State Alignment-based Imitation Learning". Under maintenance☆17May 1, 2020Updated 6 years ago