A set of RL experiments. Currently including: (1) the MDP rank experiment, based on policy gradient algorithm
☆27Feb 7, 2022Updated 4 years ago
Alternatives and similar repositories for RL
Users that are interested in RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository for SIGIR'18 paper: "Ranking for Relevance and Display Preferences in Complex Presentation Layouts"☆16Aug 28, 2018Updated 8 years ago
- A deep reinforcement learning approach to search engine ranking (PyTorch). Final Project for UC Berkeley's CS 285: Deep Reinforcement Lea…☆27May 5, 2024Updated 2 years ago
- ☆12Jun 17, 2019Updated 7 years ago
- Predict and recommend the news articles, user is most likely to click in real time.☆32Apr 3, 2018Updated 8 years ago
- The code to reproduce the experimental results for "A Text-based Deep Reinforcement Learning Framework for Interactive Recommendation".☆12Mar 18, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- C++ library to parse WARC files☆11Jan 27, 2019Updated 7 years ago
- set of utilities helping me build and navigate my personal flat-file markdown wiki☆16Apr 24, 2020Updated 6 years ago
- ☆10Apr 18, 2017Updated 9 years ago
- Learning to Recommend using a Deep Reinforcement Agent☆23Apr 2, 2017Updated 9 years ago
- Official repository of "Efficient and Effective Query Expansion for Web Search", Short Paper @ CIKM 2018☆15Nov 17, 2019Updated 6 years ago
- ☆10May 22, 2023Updated 3 years ago
- Lecture on SIMD units☆11Feb 28, 2017Updated 9 years ago
- Code for 'Diff-MSR: A Diffusion Model Enhanced Paradigm for Cold-Start Multi-Scenario Recommendation' accepted to WSDM 2024☆15Aug 1, 2025Updated last year
- Kernelized rank learning for personalized drug recommendation☆16Oct 8, 2018Updated 7 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 用强化学习来玩微信跳一跳☆11Jul 10, 2022Updated 4 years ago
- Code for "Learning Deep Features in Instrumental Variable Regression" (https://arxiv.org/abs/2010.07154)☆16Sep 16, 2024Updated last year
- Compressed Bitmap in C++ for bitmap Indexes.☆11Dec 17, 2019Updated 6 years ago
- GPU-Accelerated Faster Decoding of Integer Lists☆13Aug 20, 2019Updated 7 years ago
- ☆14Nov 21, 2023Updated 2 years ago
- ☆23Dec 31, 2020Updated 5 years ago
- Joint Optimization of Cascade Ranking Models (WSDM 19)☆13Jun 21, 2022Updated 4 years ago
- Pretraining summarization models using a corpus of nonsense☆13Sep 28, 2021Updated 4 years ago
- A RAG system is just the beginning of harnessing the power of LLM. The next step is creating an intelligent Agent. In Agentic RAG the Ag…☆14May 31, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A game search and evaluation parameter tuner using optuna framework☆14Jul 25, 2026Updated last month
- Sequential recommendation algorithm☆28Dec 28, 2018Updated 7 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- ☆18Jul 9, 2018Updated 8 years ago
- MMR for information retrieval☆18Sep 22, 2017Updated 8 years ago
- Tensorflow implementation for "Generative Adversarial User Model forReinforcement Learning Based Recommendation System"☆131Sep 10, 2019Updated 6 years ago
- 64-bit integer compression algorithms in Java☆15Nov 11, 2018Updated 7 years ago
- This program implements the following graph reordering technique: Laxman Dhulipala, Igor Kabiljo, Brian Karrer, Giuseppe Ottaviano, Serg…☆11Sep 13, 2018Updated 7 years ago
- DimmWitted Gibbs Sampler in C++ — ⚠️🚧🛑 REPO MOVED TO DEEPDIVE 👉🏿☆17Jan 23, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for paper "Hierarchically Decoupled Imitation for Morphological Transfer"☆17Mar 24, 2023Updated 3 years ago
- ☆20Sep 1, 2021Updated 5 years ago
- Implementation of the algorithm in Python 3, TensorFlow and OpenAI Gym☆176Mar 1, 2018Updated 8 years ago
- Code associated with the NeurIPS19 paper "Weighted Linear Bandits in Non-Stationary Environments"☆17Nov 14, 2019Updated 6 years ago
- A Tensorflow implementation of the Deep Listwise Context Model (DLCM) for ranking refinement.☆136Feb 1, 2023Updated 3 years ago
- Code for Policy Learning for Fairness in Ranking paper at NeurIPS 2019☆20Apr 20, 2022Updated 4 years ago
- ICLR Reproducibility Challenge for Discriminator-Actor-Critic☆20Jan 7, 2019Updated 7 years ago