A comprehensive guide to Reinforcement Learning
☆201Sep 16, 2026Updated this week
Alternatives and similar repositories for rl-handbook
Users that are interested in rl-handbook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CUAI's RL/interactive learning reading group website + materials :)☆37Aug 13, 2026Updated last month
- ☆17Apr 29, 2025Updated last year
- A blazing-fast directory jumper written in Rust☆16Jul 27, 2026Updated last month
- Repository for the paper "U-Net Transplant: The Role of Pre-training for Model Merging in 3D Medical Segmentation" accepted @ MICCAI2025☆31Jun 26, 2025Updated last year
- Multi-purpose data analysis framework based on Bayesian networks and Causal models☆28Nov 21, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Repo for focus response library☆19Oct 20, 2025Updated 11 months ago
- A modern, real-time monitoring dashboard built with FastAPI and Svelte. This application demonstrates real-time data streaming using Serv…☆19Mar 31, 2025Updated last year
- Model Context Protocol (MCP) Server for Handwriting OCR☆17Jul 22, 2026Updated last month
- LaTeX document compiler with MCP interface — AI agents author documents via structured model, TeXFlow handles LaTeX mechanics☆20Mar 5, 2026Updated 6 months ago
- A load balanced threadpool.☆31Jun 14, 2022Updated 4 years ago
- RIDE is a Python library designed to accelerate Dijkstra's algorithm on diverse graph structures using a hierarchical approach☆38Nov 21, 2025Updated 9 months ago
- Stateless LLM runtime that dynamically routes, loads, executes, and unloads models per request with bounded VRAM caching and intelligent …☆42Apr 12, 2026Updated 5 months ago
- Recent papers on Graph Neural Networks-based Recommender System.☆12Aug 21, 2023Updated 3 years ago
- ☆19Dec 3, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- MCP server for Git with local Ollama — zero tokens for git operations☆43Updated this week
- Generalizing to New Physical Systems via Context-Informed Dynamics Model☆28May 6, 2024Updated 2 years ago
- Making MoE fine-tuning accessible to anyone with consumer hardware.☆17Jul 30, 2026Updated last month
- ☆13Sep 7, 2026Updated last week
- Postgresql MCP server written in Rust☆22Jul 14, 2026Updated 2 months ago
- Official Implementation of "Decentralized Transformers with Centralized Aggregation are Sample-Efficient Multi-Agent World Models" (accep…☆21Dec 14, 2025Updated 9 months ago
- A skill for AI agents. A virtual beer and wine sommelier that learns your tastes and suggests the best choice.☆15Jul 7, 2026Updated 2 months ago
- Fleet Python SDK☆17Aug 9, 2026Updated last month
- 🤖 AI GitHub App that automatically reviews PRs, triages issues, and monitors repository health using LLMs.☆24Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Blog-style notebook indexer and renderer for org documents☆11Feb 22, 2025Updated last year
- Autoregressive Image Generation☆31Jun 13, 2025Updated last year
- win32 native frontend for llama-cli☆14Nov 2, 2024Updated last year
- Post-training framework for large models, from new objectives to new rollout systems.☆215Aug 21, 2026Updated 3 weeks ago
- finetune method to create think/model/requires tags to allow LLMs to write programs for things they can calculate instead of hallucinatin…☆15Apr 8, 2026Updated 5 months ago
- Postgres notifications listener for pg-promise☆21Nov 1, 2025Updated 10 months ago
- ReLAx - Reinforcement Learning Applications Library☆16Feb 19, 2023Updated 3 years ago
- Notes for EE364a - Convex Optimization I @ Stanford (will update Ch 6 - Ch 13 later)☆12Jul 30, 2019Updated 7 years ago
- A reinforcement learning based drone controller and simulator, backed by @PufferAI☆42May 23, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆16May 14, 2025Updated last year
- eco4cast library aims to reduce carbon footprint of machine learning models with predictive cloud computing scheduling☆16Aug 26, 2024Updated 2 years ago
- Open-source AI agent that reasons by blending many AI models, does real work on its own, and keeps learning — safe, self-hostable, and fu…☆26Updated this week
- Lightweight reverse proxy + admin UI that turns your backend endpoints into multi-tenant Model Context Protocol tools for OpenServ—or any…☆16Sep 8, 2026Updated last week
- Automatically show hidden emphasis markers at point in org mode.☆17Jul 27, 2025Updated last year
- A local tiddlywiki server☆12May 6, 2017Updated 9 years ago
- NEURON OS — personal AI assistant. Deploy waitlist: https://render.com/deploy?repo=https://github.com/yachitguliani/personal-assitant☆15Jul 2, 2026Updated 2 months ago