Exercise Solutions for Reinforcement Learning: An Introduction [2nd Edition]
☆107Sep 14, 2022Updated 3 years ago
Alternatives and similar repositories for rlai-exercises
Users that are interested in rlai-exercises are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Exercise Solutions for Reinforcement Learning: An Introduction [2nd Edition]☆156Jan 14, 2021Updated 5 years ago
- Implementations for solutions to programming exercises of Reinforcement Learning: An Introduction, Second Edition (Sutton & Barto)☆33Jun 23, 2022Updated 4 years ago
- Implementation of Reinforcement Learning algorithms in Python, based on Sutton's & Barto's Book (Ed. 2)☆159May 3, 2020Updated 6 years ago
- My Solutions to Sutton and Barto exercises, 2nd edition☆14Apr 27, 2018Updated 8 years ago
- Quantized Tensorflow Estimator for Google EdgeTpu Accelerator, Dev Board and Intel Neural Compute Stick 2☆14May 9, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Python Implementation of Reinforcement Learning: An Introduction☆14,732Aug 9, 2024Updated last year
- ☆50May 25, 2018Updated 8 years ago
- ☆27Mar 11, 2025Updated last year
- Notes and exercise solutions for second edition of Sutton & Barto's book☆409Oct 2, 2022Updated 3 years ago
- Solutions of Reinforcement Learning, An Introduction☆2,424Jul 10, 2025Updated last year
- Implementation of BIMRL: Brain Inspired Meta Reinforcement Learning - Roozbeh Razavi et al. (IROS 2022)☆10Dec 1, 2022Updated 3 years ago
- PyTorch - Implicit Quantile Networks - Quantile Regression - C51☆22Jul 26, 2019Updated 7 years ago
- Implementation of Reinforcement Learning Algorithms. Python, OpenAI Gym, Tensorflow. Exercises and Solutions to accompany Sutton's Book a…☆22,086Jul 13, 2023Updated 3 years ago
- A Keras model that addresses the Quora Question Pairs dyadic prediction task.☆14Feb 18, 2017Updated 9 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Examples of published reinforcement learning algorithms in recent literature implemented in TensorFlow☆103Aug 3, 2020Updated 5 years ago
- Code used for the master thesis at MIIS (UPF)☆16Dec 1, 2016Updated 9 years ago
- Submissions for Github Dev Demo Days☆11Feb 27, 2022Updated 4 years ago
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- Using Kubernetes for Fog Computing☆38Dec 15, 2025Updated 7 months ago
- Deep Reinforcement Learning Policy Gradients Method - Pong game - Keras☆22Jun 15, 2018Updated 8 years ago
- openAI gym env for reversi/othello game☆20Nov 6, 2023Updated 2 years ago
- Dynamics and Control of a Six-wheeled Rover with Rocker-Bogie Suspension☆14Jan 12, 2022Updated 4 years ago
- python interface to bnlearn and other probabilistic graphical model libraries☆10Mar 26, 2020Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A library for developing and applying Seldonian algorithms☆12Jan 13, 2024Updated 2 years ago
- Stanford cs231n (HKUST COMP4901J Fall 2018 Deep Learning in Computer Vision) Assignment Repository☆10Jan 29, 2019Updated 7 years ago
- Implementation of zero-velocity updates for motion tracking☆16Dec 15, 2014Updated 11 years ago
- test tcp congestion fairness on mininet☆10Aug 18, 2020Updated 5 years ago
- trading by Deep Q-Network☆15Oct 20, 2016Updated 9 years ago
- Blog post: how to do deterministic policy gradient with gumbel softmax and why you should do it.☆12Jun 20, 2017Updated 9 years ago
- ☆102May 10, 2020Updated 6 years ago
- A Hindi-English Dataset for Text Normalization☆18Jan 3, 2022Updated 4 years ago
- This project contains several Deep Reinforcement Learning method and some experiments basd on OpenAi gym.☆19Jan 28, 2018Updated 8 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- tech note☆16Mar 2, 2023Updated 3 years ago
- Slides for the tutorial talk on Bayesian Machine Learning at PyCon 2017☆10May 19, 2017Updated 9 years ago
- RC-NFQ: Regularized Convolutional Neural Fitted Q Iteration. A batch algorithm for deep reinforcement learning. Incorporates dropout regu…☆12Mar 17, 2021Updated 5 years ago
- [INACTIVE] A bunch of articficial intelligence algorithms☆11May 14, 2016Updated 10 years ago
- G-HER algorithm☆18May 24, 2019Updated 7 years ago
- Deep reinforcement learning baselines base on OpenAI. More algorithms are included, such as Rainbow: Combining Improvements in Deep Rei…☆35Aug 23, 2018Updated 7 years ago
- PowerShell によって Windows10 のキッティングに必要な全工程を自動的に完了。☆12Jun 10, 2025Updated last year