Collection of Deep Reinforcement Learning Jupyter Notebooks. Each notebook is self-contained and presents single algorithm. These include DP, MC, TD, SARSA, Q-Learning and DQNs.
☆42Mar 7, 2020Updated 6 years ago
Alternatives and similar repositories for rl-sketchpad
Users that are interested in rl-sketchpad are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Image Search Engine with HuggingFace Sentence Transformer☆12Aug 31, 2023Updated 2 years ago
- Multi-Agent LLM System for Digital Scam Protection☆15Dec 19, 2024Updated last year
- Applications of reinforcement learning to Groebner basis computation.☆14Jun 13, 2021Updated 5 years ago
- Agent Watch is an AgentOps monitoring library designed for Crew AI applications.☆23Dec 2, 2024Updated last year
- Code Repository for Blog - How to Productionize Large Language Models (LLMs)☆12Mar 27, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code repository for the paper "Learning partial differential equations for biological transport models from noisy spatiotemporal data"☆11Jul 3, 2019Updated 7 years ago
- (ICLR 2021) Learning to Represent Action Values as a Hypergraph on the Action Vertices☆23Jun 22, 2021Updated 5 years ago
- A collection of various projects related to Reinforcement Learning☆19Feb 22, 2021Updated 5 years ago
- This repository contains resources, documentation and artifacts describing LLM agents☆15Jan 22, 2025Updated last year
- In this course navigates through the LLMOps pipeline, enabling you to preprocess training data for supervised fine-tuning and deploy cust…☆15Feb 13, 2024Updated 2 years ago
- ☆14Apr 22, 2024Updated 2 years ago
- we generate captions to the images which are given by user(user input) using prompt engineering and Generative AI☆10Aug 24, 2024Updated 2 years ago
- A PyTorch Implementation of PlaNet: A Deep Planning Network for Reinforcement Learning☆13Aug 31, 2020Updated 5 years ago
- MLflow is Open source platform for the machine learning lifecycle so here you can learn MLflow End to End Example with Prediction.☆13Jun 14, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Regression in Convolutional Neural Network applied to Plant Leaf Count☆20Sep 6, 2022Updated 3 years ago
- Tutorial covering Open Source tools for Source Separation.☆15Nov 12, 2021Updated 4 years ago
- Getting Started in Imitation Learning☆13Mar 3, 2025Updated last year
- ☆13Apr 26, 2025Updated last year
- Contextual Bandits Action Elimination DQN☆21Jun 25, 2018Updated 8 years ago
- React.js Babylon.js WebGL project☆10Jan 25, 2022Updated 4 years ago
- posenet+LSTM implementation with Keras& TensorFlow☆16Nov 28, 2019Updated 6 years ago
- Tutorial on NetworkX originally given at NetsciX 2016 School of Code☆15Jul 22, 2024Updated 2 years ago
- Unity Networking Library Benchmark on Bad Network Conditions☆18Sep 1, 2025Updated 11 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ZeroMat as presented at ICISCAE 2021☆12Jun 2, 2022Updated 4 years ago
- Experiments to train transformer network to master reinforcement learning environments.☆32Mar 14, 2021Updated 5 years ago
- Official codebase for our NeurIPS paper, Symmetry-Informed Governing Equation Discovery.☆11Nov 13, 2024Updated last year
- Official repo of ICASSP 2022 paper - Don't Separate, Learn to Remix: End-to-End Neural Remixing with Joint Optimization☆20Jan 7, 2025Updated last year
- Repository of UML diagrams generated using clang-uml☆15Mar 5, 2025Updated last year
- This project is focused on the Deployment phase of machine learning. The Docker and FastAPI are used to deploy a dockerized server of tra…☆28Jan 7, 2023Updated 3 years ago
- 🎱🤞 3D Carom Billiard Simulator☆13Oct 26, 2023Updated 2 years ago
- Tensorflow implementation for "Noisy network for exploration"☆19Aug 2, 2017Updated 9 years ago
- NLP/LLM Mlops Pipeline to dev/train/evaluation, scalable deploy and monitoring systems.☆23Mar 15, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Author implementation of Monte Carlo Augmented Actor Critic in PyTorch☆18Oct 24, 2022Updated 3 years ago
- A python package to represent data using musical notes.☆12Dec 31, 2020Updated 5 years ago
- Reimplementation of FiveThirtyEight NBA Elo-only Model (Not CARMElo!)☆12Jul 6, 2019Updated 7 years ago
- Official code of Nash-DQN for paper: Nash-DQN algorithm for two-player zero-sum Markov games, details see our paper: A Deep Reinforcement…☆22Aug 26, 2022Updated 4 years ago
- Lab files of IBM's Qiskit Global Summer School 2020.☆18Sep 3, 2020Updated 5 years ago
- project based on 3d topography mapping using a drone☆13Oct 24, 2022Updated 3 years ago
- AI Agents with Google's Gemini Pro and Gemini Pro Vision Models☆29Jan 19, 2024Updated 2 years ago