Reimplementation of simple policy gradient algorithms such as REINFORCE and Actor-Critic methods.
☆17Aug 26, 2023Updated 3 years ago
Alternatives and similar repositories for pytorch_simple_policy_gradients
Users that are interested in pytorch_simple_policy_gradients are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Tensorflow implementation of SNAIL and RL2☆11Aug 17, 2019Updated 7 years ago
- PyTorch implementation of the ICML 2020 paper "Latent Bernoulli Autoencoder"☆25Apr 8, 2021Updated 5 years ago
- Awesome RL: Papers, Books, Codes, Benchmarks☆119Oct 29, 2023Updated 2 years ago
- These are my learning algorithm solutions to OpenAI Gym environments.☆11May 9, 2017Updated 9 years ago
- Variational Reinforcement Learning☆18Jul 25, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A set of Deep Reinforcement Learning Agents implemented in Tensorflow.☆13Feb 5, 2017Updated 9 years ago
- Clean, extensible implementation of MACAW [ICML 2021]☆12Dec 7, 2021Updated 4 years ago
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- Reinforcement learning approach to the prisoner's dilemma, based on Q learning☆13Dec 1, 2017Updated 8 years ago
- The Easiest Pytorch Implementation of Branching-DQN☆12Feb 10, 2021Updated 5 years ago
- A TensorFlow 2.0 with eager execution implementation of Pytorch OpenAI few-shot regression toy example☆16Jun 24, 2019Updated 7 years ago
- The official PyTorch implementation of the paper "Generalizing Consistency Policy to Visual RL with Prioritized Proximal Experience Regul…☆15Nov 10, 2024Updated last year
- E-MAML, and RL-MAML baseline implemented in Tensorflow v1☆17Dec 7, 2019Updated 6 years ago
- ☆15Mar 21, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A Reinforcement Learning Friendly Simulator for Mobile Robot☆17Jan 5, 2025Updated last year
- Self-implemented code for Model-Based Meta-Reinforcement Learning☆17Apr 28, 2019Updated 7 years ago
- A3C-LSTM algorithm tested on CartPole OpenAI Gym environment☆48Jul 4, 2018Updated 8 years ago
- Color: Train a Real-world Local Path Planner in One Hour via Partially Decoupled Reinforcement Learning and Vectorized Diversity☆23Dec 23, 2024Updated last year
- Financial Analysis and Algorithmic Trading Strategies in Python☆11Feb 16, 2023Updated 3 years ago
- Revisiting Discrete Soft Actor-Critic Accepted by Transactions on Machine Learning Research (TMLR)☆30Nov 23, 2024Updated last year
- [NeurIPS 2025] Official PyTorch implementation of "Token Bottleneck: One Token to Remember Dynamics"☆32Feb 2, 2026Updated 7 months ago
- Solution for Taxi env using HRL (Hierarchical reinforcement learning) (2018)☆21Nov 3, 2019Updated 6 years ago
- Set up a forge testing env instantly w/ ds-test, solmate + openzeppelin preinstalled.☆10Apr 14, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Bidirectional transformation between Yao IR and QASM.☆11Dec 6, 2020Updated 5 years ago
- A Foundry template to compile and test Fe contracts.☆14Apr 6, 2023Updated 3 years ago
- Limits asset outflows from contracts within customisable timeframes☆11May 7, 2022Updated 4 years ago
- Mutual Information State Intrinsic Control (ICLR 2021 Spotlight)☆39Mar 1, 2021Updated 5 years ago
- openAI gym env for reversi/othello game☆20Nov 6, 2023Updated 2 years ago
- Collection of consensus algorithm implementation for research and simulations☆12Feb 24, 2023Updated 3 years ago
- Check the dependencies of tools-deps-based Clojure projects for vulnerabilities☆13Apr 17, 2019Updated 7 years ago
- ☆14May 25, 2023Updated 3 years ago
- Permissionless pooling of NFT's into an ERC20.☆14Dec 22, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Wicked fast, thread safe in-memory key/object store for C++☆11Dec 8, 2016Updated 9 years ago
- The state-of-art deep rl algorithms for Montezuma's revenge☆28Oct 28, 2018Updated 7 years ago
- Implementation of the model from "Faster sorting algorithms discovered using deep reinforcement learning" that discovered an all-new ult…☆11Aug 29, 2023Updated 3 years ago
- ☆10Jul 21, 2019Updated 7 years ago
- 🤖 This project involves developing a Maximal Extractable Value (MEV) agent designed to optimize order execution by matching a set of ord…☆14Nov 22, 2024Updated last year
- Puppet module for managing Hashicorp's Vault☆11Jan 4, 2020Updated 6 years ago
- ☆14Mar 21, 2025Updated last year