Solutions and figures for problems from Reinforcement Learning: An Introduction Sutton&Barto
☆20Jul 16, 2019Updated 7 years ago
Alternatives and similar repositories for ReinforcementLearning_Sutton-Barto_Solutions
Users that are interested in ReinforcementLearning_Sutton-Barto_Solutions are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python implementations of the RL algorithms in examples and figures in Sutton & Barto, Reinforcement Learning: An Introduction☆100Oct 31, 2018Updated 7 years ago
- Code repository for the paper "Learning partial differential equations for biological transport models from noisy spatiotemporal data"☆11Jul 3, 2019Updated 7 years ago
- ☆33Mar 10, 2024Updated 2 years ago
- Reinforcement Learning to teach a Neato to follow a line.☆10Apr 2, 2017Updated 9 years ago
- Public Repo for the paper "Overcoming The Spectral-Bias of Neural Value Approximation"☆11May 25, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 📖Learning reinforcement learning by implementing the algorithms from reinforcement learning an introduction☆84Mar 8, 2026Updated 6 months ago
- ROS2 node that collects metrics about system resource consumption and publishes them to a topic to be emitted to CloudWatch Metrics.☆16Feb 8, 2022Updated 4 years ago
- ☆12Aug 28, 2020Updated 6 years ago
- Coursera Lesson 2: Mapping Data to Python☆16Aug 8, 2024Updated 2 years ago
- ☆11Aug 22, 2017Updated 9 years ago
- posenet+LSTM implementation with Keras& TensorFlow☆16Nov 28, 2019Updated 6 years ago
- Collision-detection and collision-avoidance navigation demonstration using a feedforward neural network.☆13Nov 4, 2018Updated 7 years ago
- MaxSum is an algorithm about Distributed Constraint Optimization Problems (DCOPs)☆11Jan 15, 2018Updated 8 years ago
- ROS-based second-generation command & control system for marine vehicles☆13Apr 15, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- EVM in python from scratch because why not☆11Aug 22, 2022Updated 4 years ago
- Adding Dreamer-v3's implementation tricks to CleanRL's PPO☆16May 19, 2023Updated 3 years ago
- A NeRF dataset☆13Jan 18, 2023Updated 3 years ago
- The core library of the DFKI multisensor pipeline framework.☆11May 23, 2022Updated 4 years ago
- ☆13Apr 25, 2024Updated 2 years ago
- Reinforcement Learning - Implementation of Exercises, algorithms from the book Sutton Barto and David silver's RL course in Python, OpenA…☆25May 3, 2020Updated 6 years ago
- A mini racetrack world for developing and testing robots with AWS RoboMaker and Gazebo simulations.☆15Sep 8, 2020Updated 6 years ago
- Simulated Model Predictive Controller (MPC) for an inverted pendulum on a cart in Python☆21Nov 4, 2020Updated 5 years ago
- Code from my Deep Learning in Motion video course from Manning Publications☆10Jul 31, 2018Updated 8 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Python MUD/MUX/MUSH/MU* development system☆26Oct 30, 2015Updated 10 years ago
- Implementations for solutions to programming exercises of Reinforcement Learning: An Introduction, Second Edition (Sutton & Barto)☆33Jun 23, 2022Updated 4 years ago
- This tutorial accompanies the NSF-CBMS Conference and Software Day on Topological Methods in Machine Learning and Artificial Intelligence…☆22May 18, 2019Updated 7 years ago
- Explore a new way to craft A.I. solutions using a modern approach to machine reasoning.☆11Sep 3, 2020Updated 6 years ago
- Provides a demo of micro-ROS based on a Kobuki and an Olimex STM32-E407 board.☆15Mar 9, 2021Updated 5 years ago
- Implementation for "Statistical arbitrage in the US equities market" by Marco Avellaneda and Jeong-hyun Lee☆28Dec 10, 2018Updated 7 years ago
- Collaborative markdown with math☆13Sep 16, 2014Updated 12 years ago
- Curso de procesamiento de imágenes con Python☆11Feb 26, 2020Updated 6 years ago
- ☆11Feb 29, 2020Updated 6 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Gazebo support for the RoboCup 3D simulation league.☆12May 3, 2020Updated 6 years ago
- Code for "Dream and Search to Control: Latent Space Planning for Continuous Control"☆12Jul 12, 2021Updated 5 years ago
- Learning to Recommend using a Deep Reinforcement Agent☆23Apr 2, 2017Updated 9 years ago
- Contextual Bandit Algorithms (+Bandit Algorithms)☆22Oct 18, 2019Updated 6 years ago
- Group project for the WorldQuant University module, risk management.☆13Feb 3, 2019Updated 7 years ago
- ☆10Jul 21, 2019Updated 7 years ago
- My personal solution for "AirLab Summer School Session 2.2"☆19Feb 17, 2022Updated 4 years ago