Distributed RL Implementation using Pytorch and Ray (ApeX(Ape-X), A3C, Distributed-PPO(DPPO), Impala)
☆27Jun 8, 2022Updated 4 years ago
Alternatives and similar repositories for DistributedRL-Pytorch-Ray
Users that are interested in DistributedRL-Pytorch-Ray are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 算法工程师技术栈学习笔记☆15Aug 22, 2022Updated 4 years ago
- This is an pytorch implementation of Distributed Proximal Policy Optimization(DPPO).☆62Jul 30, 2018Updated 8 years ago
- Implementation of Direct Preference Optimization☆17Jul 17, 2023Updated 3 years ago
- Pytorch code for "Learning Guidance Rewards with Trajectory-space Smoothing" (NeurIPS 2020)☆12Jul 7, 2021Updated 5 years ago
- A C++ Package for Solving Multiple-Phase Optimal Control Problem Using Adaptive Radau Pseudospectral Methods☆10Aug 31, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- MATLAB code of examples using Gauss pseudospectral method, MS thesis included☆10Sep 18, 2020Updated 5 years ago
- Controlled Invariant Sets in Two Moves☆14Dec 21, 2021Updated 4 years ago
- ☆40Jul 17, 2022Updated 4 years ago
- PyTorch IMPALA implementation☆27Aug 31, 2019Updated 7 years ago
- ☆10Mar 22, 2021Updated 5 years ago
- 预测-校正学习计算制导律☆13Jun 22, 2021Updated 5 years ago
- Python implementation of algorithms for multi-objective multi-agent path finding.☆12May 17, 2022Updated 4 years ago
- PyTorch Implementation of "Language as an Abstraction for Hierarchical Deep Reinforcement Learning" paper☆25Feb 14, 2022Updated 4 years ago
- ☆14Jun 13, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A PyTorch implementation of SEED, originally created by Google Research for TensorFlow 2.☆14Dec 8, 2020Updated 5 years ago
- Average-Reward Reinforcement Learning with Trust Region Methods☆11Oct 17, 2022Updated 3 years ago
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆16May 19, 2023Updated 3 years ago
- Official repo for our AAAI'21 paper, https://arxiv.org/abs/2007.12354☆30Jul 14, 2021Updated 5 years ago
- this is a reproduction of my senior's graduation project☆14Jun 21, 2022Updated 4 years ago
- ☆26Jun 14, 2022Updated 4 years ago
- ☆13Dec 12, 2021Updated 4 years ago
- Address the JSP problem through DRL, including mlp, gcn, transformer policies.☆10May 7, 2023Updated 3 years ago
- ☆13Jun 1, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 2022 WSDM 爱奇艺用户留存预测赛 第三名方案☆14Jan 29, 2022Updated 4 years ago
- Code for paper "Learning to Guide: Guidance Law Based on Deep Meta-learning and Model Predictive Path Integral Control"☆19May 26, 2019Updated 7 years ago
- deep multi agent reinforcement learning tutorial book for intermediate☆12Dec 13, 2021Updated 4 years ago
- A PyTorch Platform for Distributed RL☆755Sep 15, 2021Updated 4 years ago
- A basic Python implementation of a Legendre-Gauss-Radau pseudospectral method for computational optimal control.☆16May 9, 2024Updated 2 years ago
- ☆12Mar 15, 2022Updated 4 years ago
- Solving the CVRPTW with geatpy2☆11Mar 24, 2020Updated 6 years ago
- Material of the Computational Intelligence course which originated in University of Guilan. This is the repository of codes written in cl…☆43Dec 12, 2019Updated 6 years ago
- Project for Elective in Robotics: Control of Multi-robot system, Univ. La Sapienza Roma, 2020.☆11Jan 25, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆22May 21, 2018Updated 8 years ago
- Codes accompanying the paper "Learning Nearly Decomposable Value Functions with Communication Minimization" (ICLR 2020)☆89Dec 8, 2022Updated 3 years ago
- ☆11Jul 30, 2023Updated 3 years ago
- Bot for `Kore 2022 - Beta` https://www.kaggle.com/competitions/kore-2022-beta/overview☆10May 28, 2022Updated 4 years ago
- Cloud client for douzero training☆11Dec 26, 2021Updated 4 years ago
- Determination of optimal spacecraft landing trajectories via convex optimization☆19Jun 6, 2019Updated 7 years ago
- [RA-L + ICRA22] Learning Sparse Interaction Graphs of Partially Detected Pedestrians for Trajectory Prediction☆48Mar 19, 2022Updated 4 years ago