Policy Gradient Actor-Critic PyTorch | Lunar Lander v2
☆78May 7, 2019Updated 7 years ago
Alternatives and similar repositories for Actor-Critic-PyTorch
Users that are interested in Actor-Critic-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- OpenAI Gym's LunarLander-v2 Implementation☆42Apr 27, 2024Updated 2 years ago
- Actor Critic model to play Cartpole game☆52Aug 4, 2018Updated 8 years ago
- PyTorch implementation of DDPG algorithm for continuous action reinforcement learning problem.☆423Mar 17, 2021Updated 5 years ago
- Twin Delayed DDPG (TD3) PyTorch solution for Roboschool and Box2d environment☆107Jun 7, 2019Updated 7 years ago
- ☆10Aug 8, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2024] Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow☆44Jun 15, 2026Updated last month
- Experiments of the three PPO-Algorithms (PPO, clipped PPO, PPO with KL-penalty) proposed by John Schulman et al. on the 'Cartpole-v1' env…☆13Nov 14, 2021Updated 4 years ago
- Minimal Implementation of Deep RL Algorithms in PyTorch☆26May 10, 2020Updated 6 years ago
- This project uses gpt-4 to build agents to play one night werewolf.☆10Jul 14, 2023Updated 3 years ago
- Accepted by AROB 2021. A car-agent navigates in complex traffic conditions by Mixed_Input_PPO_CNN_LSTM model.☆14May 22, 2021Updated 5 years ago
- Minimal implementation of clipped objective Proximal Policy Optimization (PPO) in PyTorch☆2,371Jul 9, 2024Updated 2 years ago
- ☆13Jan 14, 2020Updated 6 years ago
- ☆10Jun 21, 2021Updated 5 years ago
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A library for ready-made reinforcement learning agents and reusable components for neat prototyping☆303Feb 13, 2024Updated 2 years ago
- Compact LaTeX Template for the standard institute format. This is a modification of MR Bharath's LaTeX template. I've made it more compac…☆11Dec 28, 2016Updated 9 years ago
- ☆10Sep 18, 2019Updated 6 years ago
- Quantum Principal Component Analysis (QPCA) as a generative model☆13Apr 5, 2022Updated 4 years ago
- Guardian, Reuters, Mining 크롤링 학습용 예제☆25Dec 18, 2023Updated 2 years ago
- ☆24Dec 3, 2025Updated 8 months ago
- Reimplementation of "An Object-Oriented Representation for Efficient RL"☆17Sep 12, 2024Updated last year
- A toy example of Policy Gradient implemented in Pytorch☆94Jan 24, 2018Updated 8 years ago
- Sumo OSM short usage tutorial☆15Feb 7, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for SIGKDD2025 paper: An Efficient Diffusion-based Non-Autoregressive Solver for Traveling Salesman Problem☆15Jan 28, 2025Updated last year
- 파뿌리(파이썬 뿌시는 이십대들) 강의자료☆13Nov 14, 2022Updated 3 years ago
- Bidirectionally-Coordinated Net Implements with PyTorch 1.0☆15Apr 10, 2019Updated 7 years ago
- Accompanying repository for Unsupervised Active Domain Randomization in Goal-Directed RL☆12Aug 4, 2020Updated 6 years ago
- Reinforcement Learning Benchmark☆13Sep 9, 2020Updated 5 years ago
- Reinforcement Learning Environments for Omniverse Isaac Gym☆10May 9, 2023Updated 3 years ago
- Semantic-Aware Fine-Grained Correspondence, at ECCV 2022 (Oral)☆14Oct 29, 2022Updated 3 years ago
- OpenAI LunarLander-v2 DeepRL-based solutions (DQN, DuelingDQN, D3QN)☆43Aug 11, 2021Updated 4 years ago
- TensorFlow implementation of Deep Reinforcement Learning papers☆28Dec 31, 2016Updated 9 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Solutions for different Reinforcement Learning environments☆26Aug 2, 2024Updated 2 years ago
- advantage actor-critic reinforcement learning for openai gym cartpole☆66Jul 13, 2017Updated 9 years ago
- Separating value functions across time-scales.☆18May 13, 2019Updated 7 years ago
- Codes used to perform the experiments described in this work: https://arxiv.org/abs/1904.05803☆12Aug 29, 2019Updated 6 years ago
- Suggestions for those interested in developing audio applications of machine learning☆14Jan 10, 2020Updated 6 years ago
- Thesis: Application of Reinforcement Learning for the Control of Nonlinear Dynamical Systems☆18Apr 16, 2020Updated 6 years ago
- KoRean based ELECTRA pre-trained models (KR-ELECTRA) for Tensorflow and PyTorch☆15Feb 13, 2022Updated 4 years ago