Reinforcement learning algorithm implementations and ML experimentation workspace
☆45Jun 8, 2019Updated 7 years ago
Alternatives and similar repositories for rl_implementations
Users that are interested in rl_implementations are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Mar 31, 2023Updated 3 years ago
- Deep Q-Network (DQN) to play classic Atari Games☆11Sep 18, 2017Updated 8 years ago
- A PyTorch implementation of OpenAI's REPTILE algorithm☆219Dec 31, 2019Updated 6 years ago
- Recommendation system for music.☆15Mar 12, 2023Updated 3 years ago
- ☆47Jun 19, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- InfoGAN Implementation in PyTorch☆20May 3, 2018Updated 8 years ago
- later☆10Jul 9, 2022Updated 4 years ago
- Deepmind Recurrent Environment Simulators paper implementation in tensorflow☆74Feb 2, 2018Updated 8 years ago
- Reinforcement Learning papers on exploration methods.☆19Jun 27, 2021Updated 5 years ago
- Code for the paper "On First-Order Meta-Learning Algorithms"☆1,044May 20, 2023Updated 3 years ago
- ☆33Jun 14, 2018Updated 8 years ago
- Deep reinforcement learning using an asynchronous advantage actor-critic (A3C) model.☆64Mar 10, 2018Updated 8 years ago
- Project for Course : Reinforcement Learning☆16Apr 29, 2020Updated 6 years ago
- TuffyLite is an open-source MLN inference engine that modifies the original Tuffy solver.☆27Aug 4, 2016Updated 10 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Attempt at reinforcement learning with curiosity for Sonic the Hedgehog games. Number 149 on OpenAI retro contest leaderboard, but more w…☆33Sep 17, 2018Updated 7 years ago
- A startup search engine made using embeddings built on crunchbase company descriptions☆11Dec 2, 2015Updated 10 years ago
- ☆16Feb 22, 2024Updated 2 years ago
- A gym environment for Stuart Armstrong's model of a treacherous turn.☆18Jul 28, 2018Updated 8 years ago
- GAN(TK)²: GAN Neural Tangent Kernel ToolKit☆13Jul 12, 2022Updated 4 years ago
- World Models applied to the Open AI Sonic Retro Contest☆78Jun 30, 2018Updated 8 years ago
- simple Word2vec from scratch using tensorflow for understanding☆34Dec 29, 2017Updated 8 years ago
- Introduction to mathematical programming with Pyomo (Python)☆13Oct 4, 2018Updated 7 years ago
- A python implementation of Dueling Bandit Gradient Descent (DBGD)☆24Jan 23, 2019Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Learn how to use the Ledger CLI tool and hack your finances!☆12Mar 5, 2017Updated 9 years ago
- Noisy Networks for Exploration☆187Jan 28, 2018Updated 8 years ago
- Notebook from my blog☆15Apr 9, 2017Updated 9 years ago
- Implementation of Adversarial Variational Optimization in PyTorch☆42Aug 7, 2018Updated 8 years ago
- Keras Implementation of TD3(Twin Delayed DDPG) with PER(Prioritized Experience Replay) option on OpenAI gym framework☆11May 29, 2021Updated 5 years ago
- One-shot Learning for Question-Answering in Gaokao History Challenge (COLING 2018)☆12Jul 2, 2018Updated 8 years ago
- Tensorflow Implementation of One-Shot Learning with Memory Augmented Neural Network☆48Dec 27, 2017Updated 8 years ago
- 의사결정(DP) + 강화학습(RL) + 온라인광고(OA) + 파이썬웹(Pyweb)☆10Nov 30, 2016Updated 9 years ago
- Reinforcement Learning in continuous state and action spaces. DDPG: Deep Deterministic Policy Gradient and A3C: Asynchronous Actor-Critic…☆14May 14, 2018Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- DRL-based collision avoidance for turtlebot3☆19Feb 6, 2023Updated 3 years ago
- Mobile manipulator Task and Motion Planning(TAMP) implementaion by using legacy Method (BasePlacement). This repository is Tested in C…☆19Oct 23, 2024Updated last year
- An implementation using pytorch of the models presented in the Multi-View Data Generation Without View Supervision paper.☆13Sep 30, 2019Updated 6 years ago
- Using Deep Learning and RNN/LSTM for Time Series Learning and Prediction☆17Oct 26, 2017Updated 8 years ago
- Data, classifiers, and notebooks for the LIME demonstration at NAACL 2016☆11Jun 14, 2016Updated 10 years ago
- Chatbot for the good vibes.☆11Aug 29, 2024Updated 2 years ago
- NLP tool for optimizing a resume for a job description, computing similarity, and extracting skills☆16Jun 7, 2017Updated 9 years ago