A continuous action space version of A3C LSTM in pytorch plus A3G design
☆259Oct 11, 2024Updated last year
Alternatives and similar repositories for a3c_continuous
Users that are interested in a3c_continuous are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A3C LSTM Atari with Pytorch plus A3G design☆563Apr 18, 2023Updated 3 years ago
- A high-performance Atari A3C agent in 180 lines of PyTorch☆172Jul 31, 2021Updated 4 years ago
- Trust Region Policy Optimization with TensorFlow and OpenAI Gym☆363Jun 2, 2020Updated 6 years ago
- PyTorch implementation of Asynchronous Advantage Actor Critic (A3C) from "Asynchronous Methods for Deep Reinforcement Learning".☆1,331Sep 25, 2019Updated 6 years ago
- Pytorch implementation of "FeUdal Networks for Hierarchical Reinforcement Learning" for Montezuma's Revenge☆95Jul 27, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Simple A3C implementation with pytorch + multiprocessing☆659Mar 10, 2023Updated 3 years ago
- Implementation of algorithms for continuous control (DDPG and NAF).☆313Feb 16, 2021Updated 5 years ago
- Train an RL agent to execute natural language instructions in a 3D Environment (PyTorch)☆237Apr 16, 2018Updated 8 years ago
- Evolution Strategies Tool☆958Dec 8, 2022Updated 3 years ago
- Implementation of TRPO and related algorithms☆654May 20, 2018Updated 8 years ago
- lagom: A PyTorch infrastructure for rapid prototyping of reinforcement learning algorithms.☆378Nov 19, 2022Updated 3 years ago
- ICML 2018 Self-Imitation Learning☆277Apr 18, 2020Updated 6 years ago
- PyTorch implementation of Advantage Actor Critic (A2C), Proximal Policy Optimization (PPO), Scalable trust-region method for deep reinfor…☆3,903May 29, 2022Updated 4 years ago
- Asynchronous Methods for Deep Reinforcement Learning☆588Aug 9, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implement A3C for Mujoco gym envs☆73Nov 2, 2017Updated 8 years ago
- Hybrid CPU/GPU implementation of the A3C algorithm for deep reinforcement learning.☆661Feb 25, 2020Updated 6 years ago
- Twin Delayed DDPG (TD3) PyTorch solution for Roboschool and Box2d environment☆107Jun 7, 2019Updated 7 years ago
- ☆54Feb 19, 2018Updated 8 years ago
- Reinforcement learning with unsupervised auxiliary tasks☆424Feb 13, 2019Updated 7 years ago
- ☆162Jul 21, 2017Updated 9 years ago
- 🔍 Codebase for the ICML '20 paper "Ready Policy One: World Building Through Active Learning" (arxiv: 2002.02693)☆18Jul 6, 2023Updated 3 years ago
- ☆28Oct 9, 2017Updated 8 years ago
- Highly Modular and Scalable Reinforcement Learning☆116Jan 14, 2020Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- PyTorch implementation of DDPG algorithm for continuous action reinforcement learning problem.☆422Mar 17, 2021Updated 5 years ago
- PyTorch implementation of Advantage async actor-critic Algorithms (A3C) in PyTorch☆113Apr 3, 2017Updated 9 years ago
- Rainbow: Combining Improvements in Deep Reinforcement Learning☆1,672Jan 13, 2022Updated 4 years ago
- Implementation of Meta-RL A3C algorithm☆407Feb 22, 2017Updated 9 years ago
- Gated Path Planning Networks (ICML 2018)☆180Jan 23, 2019Updated 7 years ago
- Ape-X DQN & DDPG with pytorch & tensorboard☆102Jun 18, 2019Updated 7 years ago
- An implementation of the Augmented Random Search algorithm☆433Sep 29, 2021Updated 4 years ago
- Distributed A3C☆34Dec 22, 2017Updated 8 years ago
- Deep Reinforcement Learning with pytorch & visdom☆802Jul 16, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Replicating "Asynchronous Methods for Deep Reinforcement Learning" (http://arxiv.org/abs/1602.01783)☆408Feb 25, 2017Updated 9 years ago
- Reinforcement learning framework to accelerate research☆205Aug 25, 2021Updated 4 years ago
- A collection of python Machine Learning articles and examples. You will find code related to Reinforcement Learning, Q Learning, MDP, Bel…☆191Nov 14, 2022Updated 3 years ago
- RUDDER for ATARI games with delayed rewards in OpenAI Baselines package☆268Oct 24, 2019Updated 6 years ago
- Reinforcement Learning with Deep Energy-Based Policies☆438Nov 28, 2023Updated 2 years ago
- PyTorch implementation of Deep Reinforcement Learning: Policy Gradient methods (TRPO, PPO, A2C) and Generative Adversarial Imitation Lear…☆1,285Feb 9, 2021Updated 5 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago