Minimalistic implementation of Vanilla Policy Gradient with PyTorch
☆18Jun 18, 2019Updated 7 years ago
Alternatives and similar repositories for VPG-PyTorch
Users that are interested in VPG-PyTorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pytorch implementation of Randomized Ensembled Double Q-learning (REDQ)☆21Mar 12, 2021Updated 5 years ago
- Implementation of the Self Paced Reinforcement Learning Experiments☆19Sep 27, 2023Updated 2 years ago
- ☆10Jan 21, 2021Updated 5 years ago
- Software package for intertemporal pricing optimization under reference effects and consumer heterogeneity estimation. Please see REAMDE.…☆11Mar 7, 2024Updated 2 years ago
- A multi-task deep reinforcement learning model for trading futures contracts using the Interactive Brokers API and TensorFlow☆15Feb 8, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆10Apr 24, 2021Updated 5 years ago
- My attempts at the exercises in the book: Neuronal Dynamics by Gerstner et al☆10Jan 7, 2019Updated 7 years ago
- AAC decoder for MPEG-4 and AAC files, with rodio support☆18Jul 26, 2026Updated 2 weeks ago
- Minimal PyTorch Library for Natural Evolution Strategies☆18Sep 29, 2021Updated 4 years ago
- Service Robot Simulator☆11May 3, 2020Updated 6 years ago
- Adaptable Agent Populations via a Generative Model of Policies☆12Oct 14, 2021Updated 4 years ago
- Double/Debiased Machine Learning implementation for Stata☆19Feb 12, 2026Updated 6 months ago
- This repository is Gazebo plugin for actor to navigation in simulation environment autonomously.☆11Mar 22, 2023Updated 3 years ago
- gazebo ros plugin for simulating WiFi router and receiver☆12Oct 23, 2016Updated 9 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- R package experiment☆14Apr 8, 2022Updated 4 years ago
- Landing a Spaceship using Upside-Down Reinforcement Learning (a.k.a ⅂ꓤ)☆13Oct 25, 2023Updated 2 years ago
- Paper: Challenges in High-dimensional Reinforcement Learning with Evolution Strategies☆29May 30, 2022Updated 4 years ago
- Stacking regression in Stata based on Scikit-learn☆17Jul 20, 2026Updated 3 weeks ago
- Examples for Econ 712, Fall 2013☆16Feb 17, 2020Updated 6 years ago
- ☆13Sep 28, 2021Updated 4 years ago
- Use an image segmentation to produce a RGB+D image (image + depthmap). Or use the GUI to view already-made RGB+D images in 3D, there's ev…☆18Mar 24, 2022Updated 4 years ago
- ☆19Jun 11, 2022Updated 4 years ago
- Reimplementation of simple policy gradient algorithms such as REINFORCE and Actor-Critic methods.☆17Aug 26, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Revealed preference analysis and preference estimation from choice datasets☆15May 6, 2026Updated 3 months ago
- Solution to the OpenAI Gym environment of the MountainCar through Deep Q-Learning☆24Dec 2, 2018Updated 7 years ago
- Jupyter notebooks for the pyextremes library☆15Jun 30, 2025Updated last year
- Pseudospectral Methods for Continuous-Time Heterogeneous-Agent Models☆14Aug 27, 2024Updated last year
- ☆14Nov 19, 2018Updated 7 years ago
- Self-Tuning Optimized Kalman Filtering (STOK) + DyNet simulation + connectivity metrics☆13May 10, 2021Updated 5 years ago
- Dynare Summer School 2018 material☆15Jun 22, 2018Updated 8 years ago
- Tensorflow 2.0.0 implementation of SPRT-TANDEM☆13Jun 21, 2022Updated 4 years ago
- Codebase for BRDiv: Diverse teammate generation for ad hoc teamwork☆13May 2, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Course at IIITH, M20, Probability and Statistics☆15Dec 3, 2020Updated 5 years ago
- An analysis, with a focus on demand forecasting, of transactional data associated with over 2.5 million customers and 31,868 SKUs over th…☆17Oct 4, 2020Updated 5 years ago
- This repository contains the research project that enables the robot to automatically join a group based on the modeled personal, social …☆11Nov 4, 2018Updated 7 years ago
- Proximal policy optimization in PyTorch. Easy to read and understand.☆51Oct 30, 2020Updated 5 years ago
- PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT…☆16Nov 18, 2020Updated 5 years ago
- Anime background remover☆25Sep 2, 2024Updated last year
- An API to access OpenAI Gym from other languages via Unix domain sockets☆18Mar 7, 2020Updated 6 years ago