Interaction-side integration library for Reinforcement Learning loops: Predict, Log, [Learn,] Update
☆73Mar 3, 2026Updated 5 months ago
Alternatives and similar repositories for reinforcement_learning
Users that are interested in reinforcement_learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Vowpal Wabbit examples and tutorials☆21Jan 20, 2022Updated 4 years ago
- Experimental new Python bindings for the VowpalWabbit library☆12Oct 5, 2023Updated 2 years ago
- Notes for the Neuroscience & AI Reading Course (SEM-I 2020-21) at BITS Pilani Goa Campus☆14Sep 30, 2020Updated 5 years ago
- Vowpal Wabbit is a machine learning system which pushes the frontier of machine learning with techniques such as online, hashing, allredu…☆8,701Jul 15, 2026Updated 3 weeks ago
- Repository for SAiDL Summer 2021 Induction Assignment☆21Jul 5, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- PyTorch - Implicit Quantile Networks - Quantile Regression - C51☆22Jul 26, 2019Updated 7 years ago
- ☆16Jun 5, 2017Updated 9 years ago
- This repository contains the code used in the paper Evaluating the Performance of Reinformcent Learning Algorithms☆27Aug 14, 2021Updated 5 years ago
- Boiler plate code for Torch based ML projects☆10Jul 14, 2021Updated 5 years ago
- A ROS pipeline for GPU based feature detection, description and matching☆15Jul 2, 2018Updated 8 years ago
- ☆15Jan 20, 2020Updated 6 years ago
- Neuromechanics_Course_2020☆51Jun 4, 2025Updated last year
- Reinforcement learning benchmarking.☆39Oct 22, 2018Updated 7 years ago
- A post-processing method to create fair rankings wrt ranked group fairness☆15Jun 5, 2019Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Off-policy Learning in Two-stage Recommender Systems. https://dl.acm.org/doi/pdf/10.1145/3366423.3380130☆30Jun 11, 2020Updated 6 years ago
- Source code for our paper "Top-K Contextual Bandits with Equity of Exposure" published at RecSys 2021.☆15Aug 2, 2021Updated 5 years ago
- Neural Fitted Q Iteration - First Experiences with a Data Efficient Neural Reinforcement Learning Method☆33Jul 25, 2024Updated 2 years ago
- Code for simulations in "Computational mechanisms of curiosity and goal-directed exploration"☆11May 22, 2020Updated 6 years ago
- ☆18Nov 19, 2018Updated 7 years ago
- Stochastic Gradient Riemannian Langevin Dynamics☆35May 29, 2015Updated 11 years ago
- DRIFT is a tool for Diachronic Analysis of Scientific Literature.☆126Oct 16, 2025Updated 9 months ago
- Accelerated Confergence for Counterfactual Learning to Rank☆17Jan 21, 2022Updated 4 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Simple, standalone python classes for training statistical language models using several popular smoothing methods.☆25Nov 3, 2012Updated 13 years ago
- OCaml bindings to libuv -- Cross-platform asychronous I/O☆18Jan 12, 2015Updated 11 years ago
- Context Aware Language Models☆28Jul 3, 2018Updated 8 years ago
- Python implementations of contextual bandits algorithms☆839Jun 28, 2026Updated last month
- Implementation of the X-armed Bandits algorithm, as detailed in the paper, "X-armed Bandits", Bubeck et al., 2011.☆11Jul 12, 2018Updated 8 years ago
- Using Centroids of Word Embeddings and Word Mover's Distance for Biomedical Document Retrieval in Question Answering.☆15Jul 13, 2017Updated 9 years ago
- Code in python with tensorflow of the method described in the paper Unsupervised Interpretable Pattern Discovery in Time Series Using Aut…☆21Dec 13, 2018Updated 7 years ago
- ☆12Aug 26, 2025Updated 11 months ago
- This repository contains the resources used for presentation/discussion in weekly iRE Lab meetings.☆14Sep 8, 2017Updated 8 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Introduction Notebook to Extreme Multi-Label Classification problem (XML)☆22Sep 9, 2018Updated 7 years ago
- Birkhoff decomposition for doubly stochastic matrices.☆14Sep 17, 2023Updated 2 years ago
- An LSTM based query classification for Mandrain, implemented using Tensorflow☆19Oct 5, 2016Updated 9 years ago
- Document context language models☆21Nov 13, 2015Updated 10 years ago
- Experimentation for oracle based contextual bandit algorithms.☆33Sep 12, 2022Updated 3 years ago
- Contextual Combinatorial Cascading Bandits☆10Jun 30, 2016Updated 10 years ago
- Source code for our paper "Joint Policy-Value Learning for Recommendation" published at KDD 2020.☆23Jul 6, 2023Updated 3 years ago