tensorflow implementation of Andrej Karpathy's blog about reinforcement learning. http://karpathy.github.io/2016/05/31/rl/
☆31Oct 4, 2020Updated 5 years ago
Alternatives and similar repositories for policy-gradient-pong
Users that are interested in policy-gradient-pong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- AsynchroNous Disk-based Representation of MassivE DAta: An R package aimed at replacing ff for storing large data objects.☆11Jun 11, 2026Updated 2 months ago
- Deep Reinforcement Learning Policy Gradients Method - Pong game - Keras☆22Jun 15, 2018Updated 8 years ago
- Trains an agent with (stochastic) Policy Gradients(actor-critic) on Pong. Uses OpenAI Gym.☆18Jan 10, 2025Updated last year
- State of the Art Language models and Classifier for Odia, which is spoken in the Indian state of Odisha☆14Aug 7, 2020Updated 6 years ago
- Train I3D on NTU-RGB+D dataset in keras☆11Feb 5, 2019Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆16Sep 9, 2025Updated 11 months ago
- my experimental repository of Mask R-CNN based on lightweight network☆10Apr 20, 2019Updated 7 years ago
- Oracle backend for dplyr (R package)☆14Mar 23, 2016Updated 10 years ago
- [INACTIVE] A bunch of articficial intelligence algorithms☆11May 14, 2016Updated 10 years ago
- ☆15Jan 11, 2019Updated 7 years ago
- A tool for experimenting with evolutionary optimization methods for machine learning algorithms, by distributing the workload over a larg…☆14Dec 19, 2018Updated 7 years ago
- A PyTorch implementation of Human-Level Control through Deep Reinforcement Learning☆24Jun 6, 2017Updated 9 years ago
- Automatically exported from code.google.com/p/pyrbf☆11May 4, 2015Updated 11 years ago
- Deep Q Network implements by Tensorflow☆25Mar 9, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ChatGPT Android App☆12Jun 15, 2023Updated 3 years ago
- The mechanoChemIGA code is an isogeometric analysis based code used to solve the partial differential equations describing solid mechanic…☆14Oct 15, 2020Updated 5 years ago
- coding examples to Intro to RL☆13Apr 30, 2018Updated 8 years ago
- Prune your sklearn models☆18Oct 28, 2024Updated last year
- Sequential Monte Carlo sampler for PyMC2 models.☆14Apr 4, 2018Updated 8 years ago
- Windows version of StatTag☆24Apr 14, 2026Updated 3 months ago
- Experimentation with Streamlit for personal LLM tool☆15Jun 19, 2023Updated 3 years ago
- The code used, and a docker image to run it, of the paper `Exploiting locality and physical invariants to design effective Deep Reinforce…☆13Dec 10, 2019Updated 6 years ago
- Train a quadcopter to fly with a deep reinforcement learning algorithm - DDPG☆12Jul 19, 2018Updated 8 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- a motion detector for video; written with OpenCV☆12Nov 3, 2022Updated 3 years ago
- ☆14Mar 8, 2022Updated 4 years ago
- ☆19Nov 4, 2017Updated 8 years ago
- An R interface to amCharts 4☆28Mar 15, 2023Updated 3 years ago
- An R package providing access to the OpenAI Gym API☆21Jul 1, 2017Updated 9 years ago
- Isogeometric Analysis classes for the deal.II library☆14Aug 8, 2016Updated 10 years ago
- A PyTorch implement of Dilated RNN☆11Dec 31, 2017Updated 8 years ago
- In this repository I'll be programming the cool exercises of the Book Reinforcement-Learning: An introduction by Sutton☆14Apr 15, 2018Updated 8 years ago
- Extract audio embeddings from an audio file using Python☆13Jul 25, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆14Aug 18, 2023Updated 2 years ago
- This is the repository for paper "Improving Sepsis Treatment Strategies using Deep Reinforcement Learning and Mixture-of-Experts"☆26Jul 6, 2018Updated 8 years ago
- A fast and flexible framework for data reduction in R☆37Nov 10, 2024Updated last year
- 3D topology optimzation code in C++☆19May 23, 2020Updated 6 years ago
- Comparison between Sarsa and Q-Learning algorithms on risk handling☆17Jul 10, 2017Updated 9 years ago
- Rough codebase for exploring initialization strategies for new word embeddings in pretrained LMs☆20Dec 10, 2021Updated 4 years ago
- Source code for deep learning-based reduced order models in cardiac electrophysiology. Available on doi.org/10.1371/journal.pone.0239416.☆18Sep 7, 2023Updated 2 years ago