Code associated with the NeurIPS19 paper "Weighted Linear Bandits in Non-Stationary Environments"
☆17Nov 14, 2019Updated 6 years ago
Alternatives and similar repositories for WeightedLinearBandits
Users that are interested in WeightedLinearBandits are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Study NeuralUCB and regret analysis for contextual bandit with neural decision☆103Dec 14, 2021Updated 4 years ago
- Google AI Princeton control framework☆40Nov 2, 2020Updated 5 years ago
- Useful tools and practices for Python development☆18Jul 27, 2020Updated 6 years ago
- ☆14Jun 7, 2023Updated 3 years ago
- ♊ Minimal PyTorch Twin Delayed DDPG (TD3) implementation☆10Jun 20, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Twitter follower graphs of @Die_Gruenen & @AfD, including cluster and topic analysis☆10Jul 10, 2020Updated 6 years ago
- Code for Optimistic Exploration even with a Pessimistic Initialisation☆14Aug 4, 2020Updated 6 years ago
- LibAFL 文档书 简体中文版☆17Mar 16, 2022Updated 4 years ago
- ☆12May 8, 2020Updated 6 years ago
- ☆11Oct 14, 2019Updated 6 years ago
- Welcome to FLSim_V2, a PyTorch based federated Reinforcement learning simulation framework☆10Dec 15, 2022Updated 3 years ago
- ☆10May 22, 2023Updated 3 years ago
- Companion code to CoRL 2019 paper: E Bıyık, M Palan, NC Landolfi, DP Losey, D Sadigh. "Asking Easy Questions: A User-Friendly Approach to…☆18Oct 13, 2020Updated 5 years ago
- ALNS Algorithm which optimise MINLP railroad network models (applied to Madrid's network)☆13Sep 5, 2017Updated 9 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A bottom-up model for the simulation of heat demand profiles of urban areas☆13Dec 11, 2023Updated 2 years ago
- Variational Reinforcement Learning☆18Jul 25, 2024Updated 2 years ago
- Code for 'Diff-MSR: A Diffusion Model Enhanced Paradigm for Cold-Start Multi-Scenario Recommendation' accepted to WSDM 2024☆15Aug 1, 2025Updated last year
- Performant, differentiable reinforcement learning☆23Jun 16, 2023Updated 3 years ago
- RL CIRL Research☆13Dec 8, 2022Updated 3 years ago
- Quant finance scripts☆15Apr 13, 2025Updated last year
- Code to accompany the paper "The Information Geometry of Unsupervised Reinforcement Learning"☆20Oct 6, 2021Updated 4 years ago
- Supplementary material for the paper published at ACM RecSys 2021 and its extended version accepted to ACM TORS journal☆20Jan 28, 2023Updated 3 years ago
- PyTorch training at CSCS☆22Jul 4, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 用强化学习来玩微信跳一跳☆11Jul 10, 2022Updated 4 years ago
- ☆25Feb 9, 2016Updated 10 years ago
- COBS: COmprehensive Building Simulator☆16Jun 23, 2022Updated 4 years ago
- Dockerfile that is used for the JModelica regression testing of the Buildings library and of BuildingsPy☆16Nov 22, 2023Updated 2 years ago
- Reinforcement Learning program that looks to be able to quickly learn to solve a Rubik's Cube☆15Jun 22, 2021Updated 5 years ago
- A RAG system is just the beginning of harnessing the power of LLM. The next step is creating an intelligent Agent. In Agentic RAG the Ag…☆14May 31, 2024Updated 2 years ago
- ☆16Sep 24, 2022Updated 3 years ago
- Experiments from "The Description Length of Deep Learning Models"☆10Aug 1, 2018Updated 8 years ago
- 2013 Fall Cloud Computing Project for Nerve Cloud group: MapReduce-Based Deep Learning☆15Dec 2, 2013Updated 12 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆11Dec 26, 2022Updated 3 years ago
- ☆12Aug 30, 2021Updated 5 years ago
- Code for NeurIPS 2019 paper: "Symmetry-Based Disentangled Representation Learning requires Interaction with Environments" by H. Caselles-…☆34Dec 9, 2019Updated 6 years ago
- ☆85Nov 19, 2020Updated 5 years ago
- AGAC: Adversarially Guided Actor-Critic☆47Sep 16, 2021Updated 5 years ago
- A Chainer implementation of WGAN-GP.☆12Oct 4, 2017Updated 8 years ago
- ROS package for robot learning☆17Oct 16, 2019Updated 6 years ago