A well-documented A2C written in PyTorch
☆53Jun 3, 2019Updated 7 years ago
Alternatives and similar repositories for pytorch-a2c
Users that are interested in pytorch-a2c are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of Advantage Actor-Critic (A2C)☆47Nov 25, 2017Updated 8 years ago
- Gossip-based Actor-Learner Architectures for Deep Reinforcement Learning☆20Aug 12, 2021Updated 5 years ago
- General implementation of Advantage Actor Critic using Pytorch☆28Dec 7, 2021Updated 4 years ago
- Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch☆21May 26, 2021Updated 5 years ago
- ☆14May 4, 2021Updated 5 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Simple change of a3c to a2c☆15Jun 18, 2017Updated 9 years ago
- Implementation of Model-Agnostic Meta-Learning (MAML) applied on Reinforcement Learning problems in TensorFlow 2.☆27May 11, 2021Updated 5 years ago
- PyTorch implementation of Never Give Up: Learning Directed Exploration Strategies☆57Jan 22, 2021Updated 5 years ago
- Distributed Priortized Experience Replay☆10Aug 8, 2018Updated 8 years ago
- Integrated Perception for Service RObots☆13Sep 18, 2017Updated 8 years ago
- 🌈 PyTorch Implementation for EMNLP'21 Findings "Reasoning Visual Dialog with Sparse Graph Learning and Knowledge Transfer"☆13Feb 1, 2023Updated 3 years ago
- Proximal Policy Optimization in PyTorch☆39Dec 10, 2017Updated 8 years ago
- Implementation of the two-step-task as described in "Prefrontal cortex as a meta-reinforcement learning system" and "Learning to Reinforc…☆60Mar 28, 2019Updated 7 years ago
- Delayed RL agent for non-Atari tasks, from "Acting in Delayed Environments with Non-Stationary Markov Policies", ICLR 2021.☆14Sep 12, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Study to test if Volume leak index (VLI) is a marker of severity of illness in sepsis.☆14Sep 29, 2022Updated 3 years ago
- A3C LSTM Atari with Pytorch plus A3G design☆564Apr 18, 2023Updated 3 years ago
- [CVPR 2023] Learning Geometry-aware Representations by Sketching☆16Dec 13, 2024Updated last year
- This repository contains the R code used analyse the eICU and MIMIC-III databases for the Sarkar et al paper "Performance of intensive ca…☆10Nov 27, 2020Updated 5 years ago
- Pytorch implementation of [Feudal Net](https://arxiv.org/abs/1703.01161). ([Tensorflow version](https://github.com/dmakian/feudal_networ…☆18Jun 25, 2019Updated 7 years ago
- Operationalising Sepsis-3 criteria in AmsterdamUMCdb, to accompany the AmsterdamUMCdb GitHub☆15May 20, 2024Updated 2 years ago
- ☆138Jul 25, 2024Updated 2 years ago
- Centralized cooperative reinforcement learning☆13Jan 8, 2023Updated 3 years ago
- Deep Integrated Perception framework for social service robots☆14Sep 6, 2017Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for Limbacher, T. and Legenstein, R. (2020). H-Mem: Harnessing synaptic plasticity with Hebbian Memory Networks☆16Jun 2, 2022Updated 4 years ago
- JAX implementation of GPTQ quantization algorithm☆10Jul 19, 2023Updated 3 years ago
- Implementation of Symbolic Relational Deep Reinforcement Learning based on Graph Neural Networks☆27Aug 24, 2023Updated 3 years ago
- Preprocessing and baseline for N-Omniglot. Details can be found at https://www.nature.com/articles/s41597-022-01851-z..☆19Apr 12, 2023Updated 3 years ago
- A Datasette instance for searching WebVid-10M☆15Sep 30, 2022Updated 3 years ago
- This is code for the EMNLP 2022 Paper "UniRPG: Unified Discrete Reasoning over Table and Text as Program Generation".☆10Apr 30, 2023Updated 3 years ago
- ☆30Jun 4, 2022Updated 4 years ago
- Self-Supervised Attention-Aware Reinforcement Learning☆18May 20, 2022Updated 4 years ago
- 计算中国偏度指数(CBOE skew index)☆15May 23, 2021Updated 5 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- SiDeGame - Simplified Defusal Game☆12Apr 17, 2025Updated last year
- Implementation of importance sampling, direct, and hybrid methods for off-policy evaluation.☆16Mar 28, 2020Updated 6 years ago
- BigQuery Data Connector for Dremio☆12Sep 29, 2023Updated 2 years ago
- ☆39Aug 25, 2025Updated last year
- Official implementation of "Cross-Domain Transfer via Semantic Skill Imitation", Pertsch et al., CoRL 2022☆15Dec 15, 2022Updated 3 years ago
- Active Learning in the era of Foundation Models☆14Apr 16, 2025Updated last year
- Python bindings for the Rusty Object Notation.☆20Aug 10, 2026Updated 2 weeks ago