MLP-framework (pure numpy) and DDQN-framework for OpenAI's Gym games. +test code for PPO added. +Hindsight Experience Replay(HER) bitflip-DQN example. +prioritized replay.
☆19May 24, 2018Updated 8 years ago
Alternatives and similar repositories for deep-reinforcement-learning_DDQN_PPO_HER
Users that are interested in deep-reinforcement-learning_DDQN_PPO_HER are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- First Prize Winner in HackOff-3.0 Siemens Healthineers Problem Statement number 3 on designing a Medical Chatbot.☆22Jul 11, 2023Updated 3 years ago
- A mini-MERN social website like Instagram. Deployed.☆10Jul 29, 2020Updated 6 years ago
- Implementation of Grid Search to find better hyper-parameters for decision tree to reduce the over fitting.☆12May 29, 2021Updated 5 years ago
- CATS Lab ACC data is the car-following trajectory dataset including both mix traffic and pure AV traffic.☆11Jan 6, 2023Updated 3 years ago
- Curiosity-driven Exploration by Self-supervised Prediction☆24Jun 13, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official code for Conformal Isometry of Lie Group Representation in Recurrent Network of Grid Cells (NeurIPS workshop on Symmetry and Geo…☆13Nov 1, 2022Updated 3 years ago
- A synthetic 24 hour traffic scenario for a 45 km section of the German highway A81 between Stuttgart Feuerbach - Heilbronn (Baden-Württem…☆13Oct 5, 2020Updated 5 years ago
- PyTorch implementation of the Munchausen Reinforcement Learning Algorithms M-DQN and M-IQN☆46Oct 4, 2020Updated 5 years ago
- Here is an implementation of some of a few results seen in Early Visual Concept Learning with Unsupervised Deep Learning☆28Oct 2, 2016Updated 9 years ago
- Implementation of the paper "Overcoming Exploration in Reinforcement Learning with Demonstrations" Nair et al. over the HER baselines fro…☆155Oct 25, 2021Updated 4 years ago
- Numpy implementation of Gaussian Process Regression☆10May 27, 2019Updated 7 years ago
- FLUIDS is a lightweight driving simulator for benchmarking Deep Reinforcement and Imitation learning algorithms.☆24May 3, 2019Updated 7 years ago
- 智能网联车辆和人工驾驶车辆混合行驶异质交通流特性研究☆18Sep 16, 2022Updated 3 years ago
- Author's implementation of the paper Correlated Age-of-Information Bandits.☆13Jun 19, 2021Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- code for "Data Might be Enough: Bridge Real-World Traffic Signal Control Using Offline Reinforcement Learning"☆11May 2, 2024Updated 2 years ago
- [NeurIPS 2022] Compositional Generalization in Unsupervised Compositional Representation Learning: A Study on Disentanglement and Emergen…☆13Oct 7, 2022Updated 3 years ago
- ☆33Oct 17, 2018Updated 7 years ago
- Neuroproc dataset descriptions and dictionaries☆17Jan 2, 2017Updated 9 years ago
- ☆12Dec 21, 2018Updated 7 years ago
- A pytorch implementation of "Latent Variable Dialogue Models and their Diversity"☆18Nov 30, 2017Updated 8 years ago
- Reproducing the reinforcement learning models used in "Emergence of Linguistic Communication from Referential Games with Symbolic and Pix…☆12Jun 23, 2018Updated 8 years ago
- ☆12Jan 3, 2022Updated 4 years ago
- All in AI MODELS☆13Oct 14, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Time-Contrastive Learning☆70May 30, 2018Updated 8 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- Forex Trend Finder App in React Native with Redux for Harvard CS50 Final Project☆12Dec 9, 2022Updated 3 years ago
- 2D Optical flow using NVIDIA CUDA☆19Jul 23, 2021Updated 5 years ago
- repository for my TLDR for deep learning papers (and SML papers!)☆17Jun 2, 2017Updated 9 years ago
- PaddleOCR for Chinese pdf☆11Jan 12, 2022Updated 4 years ago
- Submission for for KJSCE Hackathon 2018 [Meeting Minutes] by Team CampusConnect☆10Oct 16, 2018Updated 7 years ago
- Code repository for On the interaction between supervision and self-play in emergent communication (ICLR 2020)☆15Feb 4, 2020Updated 6 years ago
- ☆18Feb 25, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code to compute (and backpropagate) FID for a mini-batch of fake samples.☆16Oct 13, 2021Updated 4 years ago
- Catch game example is translated by TensorFlow☆16May 8, 2017Updated 9 years ago
- Pytorch implementation of the paper 'Compositional language emerge in a neural iterated learning' (ICLR 2020).☆16Oct 14, 2021Updated 4 years ago
- Python Implementation of Parameter-exploring Policy Gradients Evolution Strategy☆18Apr 2, 2020Updated 6 years ago
- Evolution Strategies in PyTorch (Tic-tac-toe)☆15Apr 22, 2017Updated 9 years ago
- Setup XDC BlockChain network in a matter of minutes.☆11Feb 2, 2019Updated 7 years ago
- Learning with latent language☆51Mar 28, 2021Updated 5 years ago