Bandit algorithms simulations for online learning
☆88May 13, 2020Updated 6 years ago
Alternatives and similar repositories for bandit_simulations
Users that are interested in bandit_simulations are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python implementations of contextual bandits algorithms☆840Jun 28, 2026Updated 2 months ago
- Yahoo! news article recommendation system by linUCB☆112Feb 1, 2018Updated 8 years ago
- ☆36Jul 8, 2019Updated 7 years ago
- Thompson Sampling for Bandits using UCB policy☆10Jul 29, 2017Updated 9 years ago
- ☆11Jun 5, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆106Sep 13, 2021Updated 5 years ago
- Contextual bandit algorithm called LinUCB / Linear Upper Confidence Bounds as proposed by Li, Langford and Schapire☆33Feb 2, 2023Updated 3 years ago
- demo of running rl-based recommender systems locally☆12Jun 11, 2022Updated 4 years ago
- Study NeuralUCB and regret analysis for contextual bandit with neural decision☆103Dec 14, 2021Updated 4 years ago
- ☆38Mar 28, 2022Updated 4 years ago
- In this notebook several classes of multi-armed bandits are implemented. This includes epsilon greedy, UCB, Linear UCB (Contextual bandit…☆92May 27, 2026Updated 4 months ago
- Implement different variants of gradient descent in python using numpy☆11Apr 23, 2019Updated 7 years ago
- Implementation of various multi-armed bandits algorithms on a 10-arm testbed.☆37Jan 16, 2020Updated 6 years ago
- Software for the experiments reported in the RecSys 2019 paper "A Simple Multi-Armed Nearest-Neighbor Bandit for Interactive Recommendati…☆21Apr 4, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15May 27, 2019Updated 7 years ago
- working example of a contextual multi-armed bandit☆55Sep 3, 2019Updated 7 years ago
- Intelligent Document Processing with AWS AI/ML, published by Packt☆12Apr 22, 2026Updated 5 months ago
- Create and revise bibtex entries from DBLP☆26Mar 3, 2026Updated 6 months ago
- A Julia Package for providing Multi Armed Bandit Experiments☆21Jul 19, 2018Updated 8 years ago
- Accompanying repository for Unsupervised Active Domain Randomization in Goal-Directed RL☆12Aug 4, 2020Updated 6 years ago
- Play with the solutions to the multi-armed-bandit problem.☆420May 21, 2024Updated 2 years ago
- Official Python implementation for paper: Probabilistic Conformal Prediction Using Conditional Random Samples☆10Jul 8, 2022Updated 4 years ago
- Network Flows Optimization - Shortest Path, Max Flow and Min Cost Flow Algorithms in Python☆11Sep 13, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Contextual bandit in python☆112Jul 7, 2021Updated 5 years ago
- Implementation of an online learning algorithm to do classification under concept drift☆23Nov 20, 2017Updated 8 years ago
- Source code for our LBR paper "Closed-Form Models for Collaborative Filtering with Side-Information" published at RecSys 2020.☆15Jul 22, 2021Updated 5 years ago
- Python library for Multi-Armed Bandits☆771Feb 11, 2020Updated 6 years ago
- Inspired by the neural style algorithm in the computer vision field, we propose a high-level language model with the aim of adapting the …☆18Nov 20, 2022Updated 3 years ago
- A custom Huggingface trainer which supports logging auxiliary losses returned by your model☆15Jul 27, 2025Updated last year
- The pytorch implementation of paper: A Graph-Enhanced Click Model for Web Search☆15Nov 17, 2021Updated 4 years ago
- [CIKM 2022] Towards Automated Over-Sampling for Imbalanced Classification☆10Mar 20, 2023Updated 3 years ago
- 关于python的面试题☆10Mar 11, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This Python code complements the video on the quantpie YouTube channel (https://www.youtube.com/c/quantpie), and contains the various fun…☆10Feb 21, 2020Updated 6 years ago
- Face Recognition System using FaceNet☆15Oct 29, 2019Updated 6 years ago
- Implementations of basic concepts dealt under the Reinforcement Learning umbrella. This project is collection of assignments in CS747: F…☆17May 21, 2018Updated 8 years ago
- A set of RL experiments. Currently including: (1) the MDP rank experiment, based on policy gradient algorithm☆27Feb 7, 2022Updated 4 years ago
- PyTorch implementation of PtrNet to solve sorting problem.☆12Dec 19, 2017Updated 8 years ago
- ☆15Jan 20, 2020Updated 6 years ago
- A deep reinforcement learning approach to search engine ranking (PyTorch). Final Project for UC Berkeley's CS 285: Deep Reinforcement Lea…☆27May 5, 2024Updated 2 years ago