☆42Aug 24, 2018Updated 7 years ago
Alternatives and similar repositories for reinforcement-learning-kdd
Users that are interested in reinforcement-learning-kdd are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- KDD Hands-On Tutorial (2018)☆29Dec 8, 2022Updated 3 years ago
- Code for "Best arm identification in multi-armed bandits with delayed feedback", AISTATS 2018.☆20Apr 3, 2018Updated 8 years ago
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- Simple but Flexible Recommendation Engine in PyTorch☆132Apr 22, 2022Updated 4 years ago
- This is the template for the gym agent.☆13Nov 8, 2018Updated 7 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆14Jul 14, 2018Updated 8 years ago
- Tutorial for PyData London 2019 on AB Test by cluster☆13Jul 12, 2019Updated 7 years ago
- micro-library to produce a couple of basic, attractive, printable plots with matplotlib☆10Mar 4, 2018Updated 8 years ago
- pyrff: Python implementation of random fourier feature approximations for gaussian processes☆29May 4, 2026Updated 3 months ago
- Review and analysis of the ICML 2017 best paper: "Understanding Black-box Predictions via Influence Functions"☆12Mar 16, 2018Updated 8 years ago
- lagom: A PyTorch infrastructure for rapid prototyping of reinforcement learning algorithms.☆378Nov 19, 2022Updated 3 years ago
- Notes for short course on econometrics in Stan☆13Jun 17, 2017Updated 9 years ago
- ☆10Jan 7, 2020Updated 6 years ago
- ☆27May 17, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Simple Semantic Segmentation☆17Sep 24, 2017Updated 8 years ago
- A simple tutorial of TensorFlow + TensorFlow / NumPy exercises☆11Feb 17, 2017Updated 9 years ago
- A Java library for preprocessing, managing and mining spatial trajectory data☆11Nov 13, 2015Updated 10 years ago
- Discussion for Stan for economists☆10Mar 29, 2016Updated 10 years ago
- Code for paper "Episodic Memory Deep Q-Networks" (https://arxiv.org/abs/1805.07603), IJCAI 2018☆63Sep 5, 2018Updated 7 years ago
- Hypothesis testing (Parametric/Non-Parametric)☆12Oct 8, 2019Updated 6 years ago
- General Latent Feature Modeling for Heterogeneous data☆50Mar 26, 2024Updated 2 years ago
- A simple murder mystery generator using Linear Logic as seen in Ceptre and MCTS-driven actors.☆12Jan 6, 2023Updated 3 years ago
- Kaggle's click through rate prediction with Spark Pipeline API☆23Feb 10, 2016Updated 10 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for 'Contrastive Multi-Document Question Generation'☆11Oct 16, 2022Updated 3 years ago
- Reward Estimation for Variance Reduction in Deep Reinforcement Learning☆11May 8, 2018Updated 8 years ago
- Autonomous exploration, active learning and human guidance with open-source Poppy humanoid robot platform and Explauto library☆18May 22, 2018Updated 8 years ago
- Some starter code for training/testing some basic CNN models given our data.☆10Feb 15, 2017Updated 9 years ago
- Convert text-intensive ICEWS data on Dataverse to conventional ISO-3166 and CAMEO codes☆14Feb 8, 2021Updated 5 years ago
- An example of how the LIME algorithm can be used to provide real-world insight into the decision processes of a 'black-box' machine learn…☆14Feb 19, 2019Updated 7 years ago
- Scientific-Computing-with-Scala_Code☆16Jan 30, 2023Updated 3 years ago
- ☆17Feb 19, 2018Updated 8 years ago
- Logarithmic Reinforcement Learning☆28Apr 7, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Scala implementations of standard algorithms for Multi-Armed Bandits Problem.☆11May 7, 2016Updated 10 years ago
- ☆11Feb 20, 2017Updated 9 years ago
- This is the LevelSpace extension repository. LevelSpace allows you to run NetLogo models |: from inside NetLogo models :|☆20Updated this week
- Monitoring Apache Kafka with Prometheus and Grafana☆10Sep 29, 2023Updated 2 years ago
- ☆25Apr 15, 2024Updated 2 years ago
- The package is developed for treatment recommendation & pairwise treatment individual effect estimation (ITE/CATE/HTE) when multiple trea…☆11Mar 9, 2023Updated 3 years ago
- Off-policy Learning in Two-stage Recommender Systems. https://dl.acm.org/doi/pdf/10.1145/3366423.3380130☆30Jun 11, 2020Updated 6 years ago