Policy Gradient Actor-Critic PyTorch | Lunar Lander v2
☆78May 7, 2019Updated 7 years ago
Alternatives and similar repositories for Actor-Critic-PyTorch
Users that are interested in Actor-Critic-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deep Q-learning approach to OpenAI Gym's Lunar Lander☆15Jul 27, 2017Updated 9 years ago
- Twin Delayed DDPG (TD3) PyTorch solution for Roboschool and Box2d environment☆105Jun 7, 2019Updated 7 years ago
- Angles Only Inital Orbit Determination from Angles only Observations☆11Feb 11, 2025Updated last year
- Tuning the PI controller parameters by using a contextual bandit approach☆15Jan 13, 2022Updated 4 years ago
- ☆10Aug 8, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Algorithms described in the paper Hindsight Credit Assignment (NeurIPS 2019).☆11Oct 27, 2019Updated 6 years ago
- [NeurIPS 2024] Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow☆44Jun 15, 2026Updated 3 months ago
- Minimal Implementation of Deep RL Algorithms in PyTorch☆26May 10, 2020Updated 6 years ago
- Scalable open-source software to run, develop, and benchmark causal discovery algorithms☆84Sep 9, 2026Updated last week
- Code for the paper "Local Causal Discovery for Estimating Causal Effects".☆12Apr 9, 2024Updated 2 years ago
- Translation and understanding of the Pop-art paper.☆18Oct 21, 2019Updated 6 years ago
- Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch☆2,384Jul 9, 2024Updated 2 years ago
- ☆13Jan 14, 2020Updated 6 years ago
- Final project for ASEN 6020 Statistical Orbit Determination. Uses Unscented Kalman Filter to estimate spacecraft state with unknown pertu…☆17Oct 11, 2018Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- A library for ready-made reinforcement learning agents and reusable components for neat prototyping☆303Feb 13, 2024Updated 2 years ago
- Compact LaTeX Template for the standard institute format. This is a modification of MR Bharath's LaTeX template. I've made it more compac…☆11Dec 28, 2016Updated 9 years ago
- Deep Reinforcement Learning by using Phasic Policy Gradient in Pytorch & Tensorflow☆20Oct 5, 2021Updated 4 years ago
- Qiskit camp 2019 hackathon: Using QAOA for solving the graph coloring problem☆11May 21, 2019Updated 7 years ago
- Quantum Principal Component Analysis (QPCA) as a generative model☆13Apr 5, 2022Updated 4 years ago
- ☆18Jun 7, 2024Updated 2 years ago
- Guardian, Reuters, Mining 크롤링 학습용 예제☆25Dec 18, 2023Updated 2 years ago
- Reimplementation of "An Object-Oriented Representation for Efficient RL"☆17Sep 12, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is pytorch version of maddpg.☆10Jun 23, 2020Updated 6 years ago
- Everything from Summer School 2022.☆17Jul 24, 2022Updated 4 years ago
- ☆21Jul 4, 2019Updated 7 years ago
- A small python package that allows the user to look up common medical abbreviations.☆13Apr 19, 2022Updated 4 years ago
- Adjustment Identification Distance: A gadjid for Causal Structure Learning☆17Apr 23, 2026Updated 4 months ago
- Sumo OSM short usage tutorial☆14Feb 7, 2018Updated 8 years ago
- Code for SIGKDD2025 paper: An Efficient Diffusion-based Non-Autoregressive Solver for Traveling Salesman Problem☆15Jan 28, 2025Updated last year
- 파뿌리(파이썬 뿌시는 이십대들) 강의자료☆13Nov 14, 2022Updated 3 years ago
- Code for paper "InfoShield: Generalizable Information-Theoretic Human-Trafficking Detection" (ICDE 2021)☆11Aug 23, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Accompanying repository for Unsupervised Active Domain Randomization in Goal-Directed RL☆12Aug 4, 2020Updated 6 years ago
- Automatic code generator for training Reinforcement Learning policies☆11Jan 3, 2021Updated 5 years ago
- Reinforcement Learning Environments for Omniverse Isaac Gym☆10May 9, 2023Updated 3 years ago
- OpenAI LunarLander-v2 DeepRL-based solutions (DQN, DuelingDQN, D3QN)☆43Aug 11, 2021Updated 5 years ago
- TensorFlow implementation of Deep Reinforcement Learning papers☆28Dec 31, 2016Updated 9 years ago
- Separating value functions across time-scales.☆18May 13, 2019Updated 7 years ago
- Codes used to perform the experiments described in this work: https://arxiv.org/abs/1904.05803☆12Aug 29, 2019Updated 7 years ago