RL-Bakery makes it easy to build production, large scale, batch Deep Reinforcement Learning applications.
☆97Oct 15, 2024Updated last year
Alternatives and similar repositories for rl-bakery
Users that are interested in rl-bakery are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆54Jul 28, 2019Updated 7 years ago
- Config files for setting up Multitenant Kubeflow on AWS with spot instances☆10Sep 15, 2020Updated 6 years ago
- Multi-objective reinforcement learning for covid-19 control☆12Aug 12, 2021Updated 5 years ago
- Non-stationary Off-policy Evaluation☆13Nov 8, 2018Updated 7 years ago
- Map maker is a command line tool and library for easily generating maps from structured data.☆16Mar 5, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Content for Udacity's Machine Learning curriculum☆10Aug 19, 2018Updated 8 years ago
- 11-785 Group Project: YouShen Poetry generation☆10Dec 23, 2020Updated 5 years ago
- Official repository for "Investigating Pre-Training Objectives for Generalization in Visual Reinforcement Learning" (ICML 2024)☆11Sep 16, 2025Updated last year
- ☆13Jun 3, 2022Updated 4 years ago
- Reinforced Recommendation toolkit built around pytorch 1.7☆586Dec 8, 2020Updated 5 years ago
- a package to support with the statistical analysis of trajectory data☆11Feb 2, 2026Updated 8 months ago
- Implementation of importance sampling, direct, and hybrid methods for off-policy evaluation.☆16Mar 28, 2020Updated 6 years ago
- Online Ranking with Multi-Armed-Bandits☆18Sep 4, 2021Updated 5 years ago
- Simulation environments for Multi-Objective Reinforcement Learning (MORL)☆17Aug 2, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Made for a reading group at the Center for Safe AGI.☆13Feb 23, 2026Updated 7 months ago
- A Multi-agent Learning Framework☆62May 10, 2021Updated 5 years ago
- Examples for variational inference☆16Jan 28, 2015Updated 11 years ago
- A Library for Modelling Probabilistic Hierarchical Graphical Models in PyTorch☆49Aug 7, 2020Updated 6 years ago
- pythorcn implementation of a vanilla RNN☆11Oct 17, 2021Updated 4 years ago
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- [NeurIPS 2025, Spotlight] An official implementation of the paper Quantization-Free Autoregressive Action Transformer☆12Mar 3, 2026Updated 7 months ago
- ☆25Nov 1, 2022Updated 3 years ago
- ☆15May 24, 2021Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Deployment strategies for AutoLFADS☆15Mar 9, 2024Updated 2 years ago
- SIR, SEIR, and beyond☆10Jul 6, 2023Updated 3 years ago
- Open-source library for a reinforcement learning research.☆53Dec 8, 2022Updated 3 years ago
- A Configurable Recommender Systems Simulation Platform☆784Jan 3, 2022Updated 4 years ago
- Detecting topic clusters in arXiv ML papers.☆14Oct 10, 2020Updated 5 years ago
- ☆14Aug 13, 2017Updated 9 years ago
- ☆10Apr 5, 2024Updated 2 years ago
- Supporting material for Princeton ORF307☆12Apr 22, 2026Updated 5 months ago
- A functional API for auction simulations☆13May 28, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A platform for Reasoning systems (Reinforcement Learning, Contextual Bandits, etc.)☆3,719Sep 1, 2026Updated last month
- ☆10Nov 4, 2019Updated 6 years ago
- ☆12Mar 6, 2020Updated 6 years ago
- TaskMet Task-driven Metric Learning for Model Learning☆21Feb 9, 2024Updated 2 years ago
- Federated Learning Infra Architecture on Kubernetes(EKS)☆20Nov 18, 2019Updated 6 years ago
- ☆12Apr 3, 2026Updated 6 months ago
- ☆12Aug 15, 2020Updated 6 years ago