Code repo for "Collapsing Bandits and Their Applications to Public Health Interventions", (NeurIPS'20)
☆11Dec 3, 2025Updated 9 months ago
Alternatives and similar repositories for collapsing_bandits
Users that are interested in collapsing_bandits are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Jupyter notebooks from our weekly (or so) hackathons☆11Dec 3, 2024Updated last year
- Density Constrained Reinforcement Learning☆12Mar 24, 2023Updated 3 years ago
- Instant search for Sphinx☆14Apr 5, 2023Updated 3 years ago
- Code for reproducing results in Delayed Impact of Fair Machine Learning (Liu et al 2018)☆15Jul 23, 2022Updated 4 years ago
- R implementation of Contextual Importance and Utility for Explainable AI☆10Updated this week
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Python library for real-time control of a robotic manipulator☆21Feb 7, 2023Updated 3 years ago
- Files related to my Summer of Science Report on Nonlinear Dynamics☆12Oct 11, 2023Updated 2 years ago
- Prosimos Simulation Engine (CLI)☆13Updated this week
- EDT: Efficient designs for Discrete Choice Experiments☆14Nov 6, 2022Updated 3 years ago
- ☆22Feb 25, 2019Updated 7 years ago
- Safe SLAC, an algorithm for safe cost-constrained reinforcement learning in high-dimensional POMDPs.☆13Mar 1, 2023Updated 3 years ago
- ☆14May 30, 2019Updated 7 years ago
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- Code and simulated data for the paper “Fair Regression for Health Care Spending”☆20Feb 21, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The Limited Multi-Label Projection Layer☆59Jul 25, 2024Updated 2 years ago
- Accompanying code for "Learning and Planning in Average-Reward Markov Decision Processes"☆15Feb 10, 2021Updated 5 years ago
- Collection of reinforcement learning algorithms☆16Oct 6, 2021Updated 4 years ago
- A minimal character-level language model using Transformer architecture in PyTorch☆10May 3, 2023Updated 3 years ago
- Non-stationary Off-policy Evaluation☆13Nov 8, 2018Updated 7 years ago
- Probabilistic planning in continuous state-action MDPs in TensorFlow.☆13Jun 21, 2022Updated 4 years ago
- Counterfactual Evaluation and Learning for Interactive Systems: Foundations, Implementations, and Recent Advances☆12Aug 14, 2022Updated 4 years ago
- Constrained episodic reinforcement learning in concave-convex and knapsack settings☆11Oct 3, 2023Updated 2 years ago
- A tour of Pomdpland☆10Aug 10, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆11Dec 27, 2021Updated 4 years ago
- reStructuredText preview in Atom using Pandoc☆10Nov 24, 2015Updated 10 years ago
- Framework for integrate BDI agents and Reinforcement Learning.☆16Aug 3, 2024Updated 2 years ago
- A Python library for logic formalisms representation and manipulation.☆16Jan 21, 2024Updated 2 years ago
- Implementation of "POPCORN: Partially Observed Prediction Constrained Reinforcement Learning" (Futoma, Hughes, Doshi-Velez, AISTATS 2020)☆11May 19, 2021Updated 5 years ago
- Code for the paper "Optimal Off-Policy Evaluation from Multiple Logging Policies"☆15Jul 17, 2021Updated 5 years ago
- Stochastic Optimization for Global Contrastive Learning without Large Mini-batches☆20Mar 31, 2023Updated 3 years ago
- R package for prostate cancer microsimulation☆21Sep 15, 2026Updated last week
- Code for Diagnosing Bottlenecks in Deep Q-learning. Contains implementations of tabular environments plus solvers.☆17May 14, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆22Jan 2, 2026Updated 8 months ago
- Convergent Policy Optimization for Safe Reinforcement Learning☆11Oct 26, 2019Updated 6 years ago
- eSNN - Learning similarity measure from data☆12Nov 28, 2019Updated 6 years ago
- A web page to collect reproduced papers in one place with their codes☆14Mar 8, 2023Updated 3 years ago
- Rust library for stochastic process mining techniques☆18Updated this week
- Code for SPIBB-DQN and Soft-SPIBB-DQN☆11May 5, 2020Updated 6 years ago
- Translates BPMN models to declarative constraints in different languages (DECLARE, SIGNAL, LTLf)☆14Nov 20, 2025Updated 10 months ago