Implementing REINFORCE algorithm on Pong, Lunar Lander and Cartplot + Medium Article
☆23Nov 24, 2020Updated 5 years ago
Alternatives and similar repositories for reinforce
Users that are interested in reinforce are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Apr 17, 2019Updated 7 years ago
- pains filter using rdktit☆11Mar 17, 2015Updated 11 years ago
- ☆13Feb 16, 2021Updated 5 years ago
- [Python] Sample Postgresql database with exercises to understand Postgresql and psycopg2 in depth.☆11Mar 17, 2018Updated 8 years ago
- ML/DL training workshops for EEE undergrads☆13Jan 16, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Stochastic Variance Reduction Policy Gradient Estimation☆11Nov 6, 2018Updated 7 years ago
- A PyTorch implement of Dilated RNN☆11Dec 31, 2017Updated 8 years ago
- Predict the steering angle of the car using CNN with center/front image as input.☆18Jan 7, 2017Updated 9 years ago
- This is the official repository for the paper "Guided Exploration with Proximal Policy Optimization using a Single Demonstration", https:…☆19Oct 5, 2021Updated 5 years ago
- Reproduction of OpenAI and DeepMind's "Deep Reinforcement Learning from Human Preferences"☆31Jul 27, 2021Updated 5 years ago
- A recurrent neural network (RNN) that generates drug-like molecules for drug discovery.☆11May 4, 2022Updated 4 years ago
- [ICANN 2022] ''An Improved Lightweight YOLOv5 Model Based on Attention Mechanism for Face Mask Detection'' Official Code☆10Feb 27, 2024Updated 2 years ago
- PyTorch tutorials.☆12Apr 6, 2021Updated 5 years ago
- [TMLR 2025] A collection of research papers on constraint inference within the field of RL☆11May 9, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- NeurIPS[2023] "Multi-Modal Inverse Constrained Reinforcement Learning from a Mixture of Demonstrations" official implement☆13Feb 19, 2024Updated 2 years ago
- Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"☆36Dec 14, 2023Updated 2 years ago
- ☆10Jan 3, 2024Updated 2 years ago
- Framework for Aerostructural Design Optimization☆11Jan 26, 2025Updated last year
- [AAMAS 2025] Privacy-preserving and Personalized RLHF, with convergence guarantees. The Code contains experiments for training multiple i…☆16Aug 31, 2026Updated last month
- Supporting codes for the numerical implementations in the paper "Operator inference for non-intrusive model reduction with quadratic mani…☆12Aug 18, 2022Updated 4 years ago
- Translating neuralese☆48Apr 26, 2017Updated 9 years ago
- 《自然语言处理——基于预训练模型的方法》全书代码实现☆12Jan 16, 2023Updated 3 years ago
- My thesis project☆10Jun 7, 2021Updated 5 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆13Sep 19, 2023Updated 3 years ago
- R2Plus1D MXNet Implementation☆11Jul 11, 2018Updated 8 years ago
- Distributed constraint satisfaction with recursive message-passing agents☆16Dec 11, 2017Updated 8 years ago
- Transport code for plasma simulations☆12Mar 27, 2026Updated 6 months ago
- This is the Pytorch implementation of paper--Training deep neural-networks using a noise adaptation layer.☆10Apr 18, 2021Updated 5 years ago
- This repository accompanies our research paper titled "An LLM-based Recommender System Environment".☆16Jul 15, 2024Updated 2 years ago
- Variational Autoencoder (VAE)-like neural network to solve ideal MHD equilibrium in a tokamak☆11May 20, 2022Updated 4 years ago
- Implementation of All ▲lgorithms in Fortran Programming Language☆16Oct 9, 2021Updated 5 years ago
- Are you you? 🔎 ML model in Python to determine if it is me who is using my computer☆18Jun 1, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆11Dec 22, 2023Updated 2 years ago
- A force field for the simulation of inorganic-organic interfaces (INTERFACE-CHARMM, INTERFACE-PCFF)☆25Feb 14, 2024Updated 2 years ago
- In this work, we present a novel approach that combines the power of Koopman operators and deep neural networks to generate a linear rep…☆13Dec 1, 2025Updated 10 months ago
- [ICLR 2025] "Understanding Constraint Inference in Safety-Critical Inverse Reinforcement Learning"☆16Nov 30, 2025Updated 10 months ago
- The implement of the policy gradient RL algorithm with pytorch☆41Dec 7, 2020Updated 5 years ago
- ☆18Sep 25, 2025Updated last year
- ☆18Nov 22, 2023Updated 2 years ago