A well-documented A2C written in PyTorch
☆53Jun 3, 2019Updated 7 years ago
Alternatives and similar repositories for pytorch-a2c
Users that are interested in pytorch-a2c are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of Advantage Actor-Critic (A2C)☆47Nov 25, 2017Updated 8 years ago
- Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch☆21May 26, 2021Updated 5 years ago
- ☆14May 4, 2021Updated 5 years ago
- PyTorch - Implicit Quantile Networks - Quantile Regression - C51☆22Jul 26, 2019Updated 7 years ago
- Implementation of Model-Agnostic Meta-Learning (MAML) applied on Reinforcement Learning problems in TensorFlow 2.☆27May 11, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyTorch implementation of Never Give Up: Learning Directed Exploration Strategies☆57Jan 22, 2021Updated 5 years ago
- Implement IMPALA architecture from Distributed Deep-RL Paper.☆15Oct 18, 2018Updated 7 years ago
- Distributed Priortized Experience Replay☆10Aug 8, 2018Updated 8 years ago
- Gym environment of simple microgrid simulation for Reinforcement Learning☆10Oct 12, 2022Updated 3 years ago
- Integrated Perception for Service RObots☆13Sep 18, 2017Updated 9 years ago
- Proximal Policy Optimization in PyTorch☆39Dec 10, 2017Updated 8 years ago
- Reinforcement learning with VizDoom platform☆12Apr 18, 2022Updated 4 years ago
- Implementation of the two-step-task as described in "Prefrontal cortex as a meta-reinforcement learning system" and "Learning to Reinforc…☆60Mar 28, 2019Updated 7 years ago
- Delayed RL agent for non-Atari tasks, from "Acting in Delayed Environments with Non-Stationary Markov Policies", ICLR 2021.☆14Sep 12, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [CVPR 2023] Learning Geometry-aware Representations by Sketching☆16Dec 13, 2024Updated last year
- ☆139Jul 25, 2024Updated 2 years ago
- Deep Integrated Perception framework for social service robots☆14Sep 6, 2017Updated 9 years ago
- JAX implementation of GPTQ quantization algorithm☆10Jul 19, 2023Updated 3 years ago
- ☆17Aug 4, 2026Updated last month
- Implementation of Symbolic Relational Deep Reinforcement Learning based on Graph Neural Networks☆27Aug 24, 2023Updated 3 years ago
- RL Algorithms☆13Mar 19, 2023Updated 3 years ago
- This is code for the EMNLP 2022 Paper "UniRPG: Unified Discrete Reasoning over Table and Text as Program Generation".☆10Apr 30, 2023Updated 3 years ago
- ☆30Jun 4, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code to reproduce the results of "Curiosity Driven Exploration of Learned Disentangled Goal Spaces"☆19Oct 26, 2018Updated 7 years ago
- Self-Supervised Attention-Aware Reinforcement Learning☆18May 20, 2022Updated 4 years ago
- ☆15Dec 3, 2022Updated 3 years ago
- 计算中国偏度指数(CBOE skew index)☆15May 23, 2021Updated 5 years ago
- ☆39Aug 25, 2025Updated last year
- Official implementation of "Cross-Domain Transfer via Semantic Skill Imitation", Pertsch et al., CoRL 2022☆15Dec 15, 2022Updated 3 years ago
- Caclualtes Bitcoin Dominance using Coingecko's API and writes it to a CSV along with the date☆10May 12, 2021Updated 5 years ago
- Active Learning in the era of Foundation Models☆14Apr 16, 2025Updated last year
- Генератор российских автомобильных номеров☆17Dec 24, 2020Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆22Oct 4, 2019Updated 6 years ago
- A short conceptual replication of "Prefrontal cortex as a meta-reinforcement learning system" in Jax.☆20Feb 27, 2023Updated 3 years ago
- Cloak + WireGuard + Docker + LAN gateway/proxy (client-side)☆15Dec 31, 2023Updated 2 years ago
- A PyTorch implementation of "Generating Sentences from a Continuous Space"☆13Feb 22, 2018Updated 8 years ago
- Pytorch implementations of RL algorithms, focusing on model-based, lifelong, reset-free, and offline algorithms. Official codebase for Re…☆109Jan 23, 2022Updated 4 years ago
- LuaJIT and luarocks in one location☆16Aug 10, 2015Updated 11 years ago
- ☆18Jul 6, 2026Updated 2 months ago