Code associated with the NeurIPS19 paper "Weighted Linear Bandits in Non-Stationary Environments"
☆17Nov 14, 2019Updated 6 years ago
Alternatives and similar repositories for WeightedLinearBandits
Users that are interested in WeightedLinearBandits are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Study NeuralUCB and regret analysis for contextual bandit with neural decision☆103Dec 14, 2021Updated 4 years ago
- Google AI Princeton control framework☆39Nov 2, 2020Updated 5 years ago
- Useful tools and practices for Python development☆18Jul 27, 2020Updated 6 years ago
- Multi-Armed Bandit algorithms applied to the MovieLens 20M dataset☆57Aug 9, 2020Updated 6 years ago
- ☆14Jun 7, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Implementation of proximal policy optimization(PPO) with tensorflow☆35Feb 10, 2018Updated 8 years ago
- Twitter follower graphs of @Die_Gruenen & @AfD, including cluster and topic analysis☆10Jul 10, 2020Updated 6 years ago
- Online multiclass boosting algorithm that uses VFDT as weak learners☆17Oct 24, 2018Updated 7 years ago
- ☆12May 8, 2020Updated 6 years ago
- ☆11Oct 14, 2019Updated 6 years ago
- ☆10May 22, 2023Updated 3 years ago
- Companion code to CoRL 2019 paper: E Bıyık, M Palan, NC Landolfi, DP Losey, D Sadigh. "Asking Easy Questions: A User-Friendly Approach to…☆18Oct 13, 2020Updated 5 years ago
- ALNS Algorithm which optimise MINLP railroad network models (applied to Madrid's network)☆13Sep 5, 2017Updated 8 years ago
- Variational Reinforcement Learning☆18Jul 25, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- (WSDM'24) Cross-modal Self-Supervised Learning for Time-series through Latent Masking☆20Feb 20, 2024Updated 2 years ago
- Avoiding catastrophic failures in reinforcement learning by learning to shape rewards.☆10Nov 13, 2017Updated 8 years ago
- Quant finance scripts☆15Apr 13, 2025Updated last year
- PyTorch training at CSCS☆22Jul 4, 2025Updated last year
- 🖖 Menlo's back-end: A system for caching blockchain data in the cloud for speed & performance☆22Jan 10, 2019Updated 7 years ago
- The Solidity based Immutable Ecosystem☆12Feb 29, 2024Updated 2 years ago
- ☆25Feb 9, 2016Updated 10 years ago
- COBS: COmprehensive Building Simulator☆16Jun 23, 2022Updated 4 years ago
- Dockerfile that is used for the JModelica regression testing of the Buildings library and of BuildingsPy☆16Nov 22, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for demonstration example-task in RUDDER blog☆24May 19, 2020Updated 6 years ago
- A RAG system is just the beginning of harnessing the power of LLM. The next step is creating an intelligent Agent. In Agentic RAG the Ag…☆14May 31, 2024Updated 2 years ago
- The repository reflects the collection of unusual findings from my smart contract audits.☆15May 29, 2024Updated 2 years ago
- 🔬 Research Framework for Single and Multi-Players 🎰 Multi-Arms Bandits (MAB) Algorithms, implementing all the state-of-the-art algorith…☆424Jun 19, 2026Updated last month
- 2013 Fall Cloud Computing Project for Nerve Cloud group: MapReduce-Based Deep Learning☆15Dec 2, 2013Updated 12 years ago
- ☆11Dec 26, 2022Updated 3 years ago
- Active Learning with Partial Feedback, ICLR 2019☆11Apr 27, 2020Updated 6 years ago
- ☆12Aug 30, 2021Updated 4 years ago
- ☆85Nov 19, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Robust policy search algorithms which train on model ensembles☆31Oct 26, 2016Updated 9 years ago
- Robustness via Retrying: Closed-Loop Robotic Manipulation with Self-Supervised Learning☆16Nov 7, 2018Updated 7 years ago
- AGAC: Adversarially Guided Actor-Critic☆47Sep 16, 2021Updated 4 years ago
- ROS package for robot learning☆17Oct 16, 2019Updated 6 years ago
- This is the source code for SDDObench: A Benchmark for Streaming Data-Driven Optimization with Concept Drift.☆25May 28, 2025Updated last year
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- The implementation of our SIGIR 2020 paper "CATN: Cross-Domain Recommendation for Cold-Start Users via Aspect Transfer Network“, Cheng Z…☆17Nov 19, 2020Updated 5 years ago