MLP-framework (pure numpy) and DDQN-framework for OpenAI's Gym games. +test code for PPO added. +Hindsight Experience Replay(HER) bitflip-DQN example. +prioritized replay.
☆19May 24, 2018Updated 8 years ago
Alternatives and similar repositories for deep-reinforcement-learning_DDQN_PPO_HER
Users that are interested in deep-reinforcement-learning_DDQN_PPO_HER are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A mini-MERN social website like Instagram. Deployed.☆10Jul 29, 2020Updated 6 years ago
- (Personal experiment) Unsupervised Predictive Memory in a Goal-Directed Agent https://arxiv.org/abs/1803.10760☆25May 3, 2019Updated 7 years ago
- Navigation agent with Bayesian relational memory in the House3D environment☆30Sep 13, 2019Updated 6 years ago
- python, ccxt, backtrader, dash☆10Apr 20, 2018Updated 8 years ago
- Implementation of Grid Search to find better hyper-parameters for decision tree to reduce the over fitting.☆12May 29, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Curiosity-driven Exploration by Self-supervised Prediction☆24Jun 13, 2019Updated 7 years ago
- Official code for Conformal Isometry of Lie Group Representation in Recurrent Network of Grid Cells (NeurIPS workshop on Symmetry and Geo…☆13Nov 1, 2022Updated 3 years ago
- ☆12Jul 13, 2023Updated 3 years ago
- ☆13Jan 16, 2018Updated 8 years ago
- PyTorch implementation of the Munchausen Reinforcement Learning Algorithms M-DQN and M-IQN☆46Oct 4, 2020Updated 5 years ago
- Here is an implementation of some of a few results seen in Early Visual Concept Learning with Unsupervised Deep Learning☆28Oct 2, 2016Updated 9 years ago
- Master's Degree final thesis project: reduce emergency vehicles travel time using V2V communications in VEINS simulator☆11Jan 9, 2022Updated 4 years ago
- Numpy implementation of Gaussian Process Regression☆11May 27, 2019Updated 7 years ago
- FLUIDS is a lightweight driving simulator for benchmarking Deep Reinforcement and Imitation learning algorithms.☆24May 3, 2019Updated 7 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Author's implementation of the paper Correlated Age-of-Information Bandits.☆13Jun 19, 2021Updated 5 years ago
- Minimal TensorFlow implementation of the Advantage Actor-Critic model for Atari games☆12Feb 15, 2018Updated 8 years ago
- ☆12Dec 14, 2021Updated 4 years ago
- Using very few experiments to efficiently learn an approximate model of an n-qubit quantum process.☆17Apr 15, 2023Updated 3 years ago
- Opensource embedded controller firmware for sipeed boards.☆14Apr 15, 2019Updated 7 years ago
- PPO with Hindsight Experience Replay (HER)☆12May 8, 2018Updated 8 years ago
- A thorough, straightforward, un-intimidating introduction to Gaussian processes in NumPy.☆16Jun 12, 2018Updated 8 years ago
- [NeurIPS 2022] Compositional Generalization in Unsupervised Compositional Representation Learning: A Study on Disentanglement and Emergen…☆13Oct 7, 2022Updated 3 years ago
- ♕ A web based and Deep-Reinforcement-Learning-powered open source chess game.☆17Feb 22, 2026Updated 5 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆33Oct 17, 2018Updated 7 years ago
- Python tool allowing easy book downloads from the terminal☆12Mar 15, 2023Updated 3 years ago
- Neuroproc dataset descriptions and dictionaries☆16Jan 2, 2017Updated 9 years ago
- Using RL-controlled vehicles as traffic regulator to reduce the travel time of emergency vehicles near intersections☆11Jan 27, 2022Updated 4 years ago
- A pytorch implementation of "Latent Variable Dialogue Models and their Diversity"☆18Nov 30, 2017Updated 8 years ago
- CS277 Project: Deep Reinforcement Learning in portfolio Management. This repo is the DQN part which implements a trading agent based on t…☆14Jan 19, 2020Updated 6 years ago
- Reproducing the reinforcement learning models used in "Emergence of Linguistic Communication from Referential Games with Symbolic and Pix…☆12Jun 23, 2018Updated 8 years ago
- All in AI MODELS☆13Oct 14, 2023Updated 2 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Forex Trend Finder App in React Native with Redux for Harvard CS50 Final Project☆12Dec 9, 2022Updated 3 years ago
- 2D Optical flow using NVIDIA CUDA☆19Jul 23, 2021Updated 5 years ago
- repository for my TLDR for deep learning papers (and SML papers!)☆17Jun 2, 2017Updated 9 years ago
- Accepted by AROB 2021. A car-agent navigates in complex traffic conditions by Mixed_Input_PPO_CNN_LSTM model.☆14May 22, 2021Updated 5 years ago
- A PyTorch Toolbox for Deep Reinforcement Learning☆10Jun 25, 2020Updated 6 years ago
- Submission for for KJSCE Hackathon 2018 [Meeting Minutes] by Team CampusConnect☆10Oct 16, 2018Updated 7 years ago
- Program to import raw brainwaves, and using FFT and Frequency Index calculate various bands of brainwaves.☆12Nov 7, 2016Updated 9 years ago