Multi-Armed Bandit Algorithms Library (MAB)
☆136Apr 13, 2026Updated 4 months ago
Alternatives and similar repositories for mabalgs
Users that are interested in mabalgs are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python implementations of contextual bandits algorithms☆839Jun 28, 2026Updated last month
- Multi-armed bandit algorithm with tensorflow and 11 policies☆16Dec 27, 2022Updated 3 years ago
- 🔬 Research Framework for Single and Multi-Players 🎰 Multi-Arms Bandits (MAB) Algorithms, implementing all the state-of-the-art algorith…☆424Jun 19, 2026Updated last month
- ☆37Jul 8, 2019Updated 7 years ago
- Python library for Multi-Armed Bandits☆771Feb 11, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- More about the exploration-exploitation tradeoff with harder bandits☆25May 12, 2019Updated 7 years ago
- Predict and recommend the news articles, user is most likely to click in real time.☆32Apr 3, 2018Updated 8 years ago
- Online Ranking with Multi-Armed-Bandits☆19Sep 4, 2021Updated 4 years ago
- A lightweight contextual bandit & reinforcement learning library designed to be used in production Python services.☆71Jun 4, 2021Updated 5 years ago
- ☆368Aug 12, 2020Updated 6 years ago
- Implementation of importance sampling, direct, and hybrid methods for off-policy evaluation.☆16Mar 28, 2020Updated 6 years ago
- Library of contextual bandits algorithms☆343Mar 14, 2024Updated 2 years ago
- Java implementation of Thompson sampling to solve the multi-armed bandit problem☆31Jun 14, 2023Updated 3 years ago
- Study NeuralUCB and regret analysis for contextual bandit with neural decision☆103Dec 14, 2021Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Python code for the post "Adversarial Bandits and the Exp3 Algorithm"☆51Jun 9, 2020Updated 6 years ago
- Contextual bandit algorithm called LinUCB / Linear Upper Confidence Bounds as proposed by Li, Langford and Schapire☆33Feb 2, 2023Updated 3 years ago
- ☆12Sep 3, 2018Updated 7 years ago
- Determinantal point processes for basket recommendations☆16Jan 15, 2019Updated 7 years ago
- A Baby Robot's Guide to Reinforcement Learning☆196Aug 26, 2023Updated 2 years ago
- Epsilon-greedy, softmax and LinUCB contextual bandit implementations [recommender systems]☆50Mar 15, 2019Updated 7 years ago
- ☆24Aug 19, 2020Updated 5 years ago
- Bandit algorithms simulations for online learning☆88May 13, 2020Updated 6 years ago
- Code associated with the NeurIPS19 paper "Weighted Linear Bandits in Non-Stationary Environments"☆17Nov 14, 2019Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Source code for our paper "Joint Policy-Value Learning for Recommendation" published at KDD 2020.☆23Jul 6, 2023Updated 3 years ago
- Accompanying repository for Unsupervised Active Domain Randomization in Goal-Directed RL☆12Aug 4, 2020Updated 6 years ago
- Code for "Best arm identification in multi-armed bandits with delayed feedback", AISTATS 2018.☆20Apr 3, 2018Updated 8 years ago
- This project is created for the simulations of the paper: [Wang2021] Wenbo Wang, Amir Leshem, Dusit Niyato and Zhu Han, "Decentralized L…☆33Oct 12, 2021Updated 4 years ago
- Implementing LinUCB and HybridLinUCB in Python.☆49May 15, 2018Updated 8 years ago
- Code for my book on Multi-Armed Bandit Algorithms☆923Jan 9, 2020Updated 6 years ago
- Reinforced Recommendation toolkit built around pytorch 1.7☆587Dec 8, 2020Updated 5 years ago
- python openflow library☆14Oct 17, 2018Updated 7 years ago
- This repository contains the code used to run generate the data splits, run the hyperparameter tunings, and export the results presented …☆14Jul 22, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Publication of the code we used in the RecSys Challenge 2018.☆12Jul 11, 2018Updated 8 years ago
- Contextual Bandits in R - simulation and evaluation of Multi-Armed Bandit Policies☆81Jul 25, 2020Updated 6 years ago
- Code for "Approaching Deep Learning through the Spectral Dynamics of Weights"☆13Oct 30, 2024Updated last year
- The code to simulate spiking neural networks as used in the paper "Spiking Time-Dependent Plasticity Leads to Efficient Coding of Predict…☆10Nov 24, 2019Updated 6 years ago
- Public repository for the work on bandit problems☆24Apr 4, 2024Updated 2 years ago
- Offline evaluation of multi-armed bandit algorithms☆23Dec 1, 2020Updated 5 years ago
- Contextual bandit in python☆112Jul 7, 2021Updated 5 years ago