ppo-lstm-parallel
☆49Mar 26, 2019Updated 7 years ago
Alternatives and similar repositories for ppo-lstm-parallel
Users that are interested in ppo-lstm-parallel are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Baseline implementation of recurrent PPO using truncated BPTT☆161Apr 28, 2024Updated 2 years ago
- ☆10Dec 10, 2021Updated 4 years ago
- An implementation of the A3C deep reinforcement learning method using a LSTM layer. Created with Tensorflow.☆29Oct 18, 2017Updated 8 years ago
- A simple RNN meta-learner☆10Dec 17, 2018Updated 7 years ago
- Solutions for different Reinforcement Learning environments☆26Aug 2, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Agile Quadruped Locomotion RL based on bullet☆23Jun 23, 2022Updated 4 years ago
- Pytorch Implementation for First Order Constrained Optimization in Policy Space (FOCOPS).☆29Dec 9, 2021Updated 4 years ago
- Tensorflow implementation of proximal policy optimization (PPO) algorithm☆13Feb 28, 2018Updated 8 years ago
- Accelerated Methods for Deep Reinforcement Learning☆49Mar 20, 2019Updated 7 years ago
- Self-implemented code for Model-Based Meta-Reinforcement Learning☆17Apr 28, 2019Updated 7 years ago
- ☆11Apr 21, 2022Updated 4 years ago
- ☆25Nov 1, 2022Updated 3 years ago
- Reinforcement learning algorithms with Generalized Advantage Estimation☆22Jun 6, 2018Updated 8 years ago
- Pytorch code for Arxiv Paper: Learning to learn: Meta-Critic Networks for Sample-Efficient Learning☆57Apr 3, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Jun 1, 2020Updated 6 years ago
- This is a robot simulation project based on Webots and C++ @ June 2020☆13Jun 24, 2020Updated 6 years ago
- Code for training policies based on paper Coordinated Multi-Agent Imitation Learning☆26Aug 7, 2017Updated 8 years ago
- 硕士毕业论文代码 深度强化学习☆10Apr 4, 2020Updated 6 years ago
- ☆12Jan 3, 2022Updated 4 years ago
- ☆23Apr 2, 2024Updated 2 years ago
- Implementation of MPC-PEARL☆13Jun 28, 2024Updated 2 years ago
- The PyTorch Implementation based on YOLOv4 of the paper: "Complex-YOLO: Real-time 3D Object Detection on Point Clouds"☆10Jan 15, 2021Updated 5 years ago
- A3C-LSTM algorithm tested on CartPole OpenAI Gym environment☆48Jul 4, 2018Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Deep Reinforcement Learning by using an on-policy adaptation of Maximum a Posteriori Policy Optimization (MPO)☆16Oct 23, 2021Updated 4 years ago
- E-MAML, and RL-MAML baseline implemented in Tensorflow v1☆17Dec 7, 2019Updated 6 years ago
- simple code to reinforcement learning☆19Aug 30, 2020Updated 5 years ago
- Foot End Trajectory of Quadruped Robot☆12Jun 8, 2020Updated 6 years ago
- JAX implementations of various deep reinforcement learning algorithms.☆25Feb 2, 2025Updated last year
- an implement of unitree_h1 robot with humanoid-gym☆17Aug 15, 2024Updated last year
- ☆10Sep 20, 2018Updated 7 years ago
- Implementation for ICML 2019 paper, EMI: Exploration with Mutual Information.☆37Dec 7, 2020Updated 5 years ago
- Recurrent continuous reinforcement learning algorithms implemented in Pytorch.☆52May 26, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Adaptation of DQN, DDQN and COMA for multi-agent Gym environments☆10Oct 3, 2023Updated 2 years ago
- ROS & Gazebo project for 1/10th scale self-driving race cars☆19Nov 12, 2020Updated 5 years ago
- ☆17Nov 16, 2022Updated 3 years ago
- Reinforcement learning notebooks☆10Apr 29, 2018Updated 8 years ago
- 基于分层强化学习和逆向强化学习的自适应巡航算法☆26Oct 8, 2019Updated 6 years ago
- ☆15Apr 5, 2023Updated 3 years ago
- Distributed Multi-Agent Cooperation Algorithm based on MADDPG with prioritized batch data.☆108Dec 6, 2020Updated 5 years ago