Repository of notes, code and notebooks in Python for the book "Reinforcement Learning: An Introduction" by Richard S. Sutton and Andrew G. Barto
☆40Mar 6, 2026Updated 5 months ago
Alternatives and similar repositories for reinforcement-learning
Users that are interested in reinforcement-learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Jupyter notebook exercises of "Reinforcement Learning: An Introduction", Richard S. Sutton and Andrew G. Barto☆27Jul 23, 2024Updated 2 years ago
- Predicting for Customers, whether they will buy car insurance or not.☆11Jan 29, 2021Updated 5 years ago
- Repository for the 2024 Telluride Topic Area Neuromorphic Systems for Space☆13Jun 21, 2024Updated 2 years ago
- nn2FPGA converts ONNX models into FPGA dataflow accelerators with seamless ONNX Runtime integration.☆22Jul 21, 2026Updated last month
- Repository of notes, code and notebooks in Python for the book Pattern Recognition and Machine Learning by Christopher Bishop☆2,627Jul 25, 2022Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Score-based Diffusion models in JAX.☆19Dec 29, 2025Updated 8 months ago
- PyTorch helper module to translate to and from NIR☆24Aug 13, 2026Updated 2 weeks ago
- We open-source our layout level fast EM simulation tool, EMSim, to the public.☆15Feb 8, 2024Updated 2 years ago
- ☆45May 4, 2025Updated last year
- Create Custom GYM Environment for SUMO and reinforcement learning agant☆15May 5, 2023Updated 3 years ago
- Format your bibtex (.bib) file to help standardize citations for conference and journal submissions☆14Nov 23, 2025Updated 9 months ago
- RDF -to- text generator, using GANs and reinforcement learning. For Google summer of code 2020.☆14Mar 25, 2023Updated 3 years ago
- End-to-End Autonomous Driving with Spiking Neural Networks☆91Jan 13, 2025Updated last year
- Implementation of Proximal Policy Optimization (PPO) for continuous action space (`Pendulum-v1` from gym) using tensorflow2.x and pytorch…☆12Aug 8, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- FMCW LiDAR implementation in CARLA simulator☆19Mar 18, 2024Updated 2 years ago
- Providing the answer to "How to do patching on all available SAEs on GPT-2?". It is an official repository of the implementation of the p…☆14Jan 26, 2025Updated last year
- This is the implementation of CounterCurate, the data curation pipeline of both physical and semantic counterfactual image-caption pairs.☆19Jun 27, 2024Updated 2 years ago
- A Deep-Reinforcement-Learning-Based Scheduler for FPGA HLS☆15Feb 27, 2021Updated 5 years ago
- Recent papers on Graph Neural Networks-based Recommender System.☆12Aug 21, 2023Updated 3 years ago
- ☆12Mar 1, 2022Updated 4 years ago
- Adaptive Machine Learning-Based Stock Prediction using Financial Time Series Technical Indicators☆10Dec 21, 2019Updated 6 years ago
- Learning ReLU INRs with B-spline wavelets.☆14Jun 5, 2024Updated 2 years ago
- My Python Intel 4004 Emulator☆19Jan 29, 2016Updated 10 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for the paper "Large Language Models Share Representations of Latent Grammatical Concepts Across Typologically Diverse Languages" (N…☆17Apr 13, 2025Updated last year
- ☆15Oct 21, 2023Updated 2 years ago
- ☆11Apr 28, 2024Updated 2 years ago
- A Deep Reinforcement Learning model for high volume and frequency Forex Portfolio Management☆13Jan 11, 2023Updated 3 years ago
- A framework for creating your own reinforcement learning environments using pybullet☆21Oct 7, 2019Updated 6 years ago
- Metrics for spiking neural networks based on torchmetrics☆13Mar 27, 2023Updated 3 years ago
- ☆10Jan 23, 2025Updated last year
- gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI☆17Jan 12, 2026Updated 7 months ago
- 10606 Fall 2023☆14Oct 13, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A fully modular framework for modeling and optimizing analog neural networks☆22Jan 19, 2026Updated 7 months ago
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- ncnn的Rust实现,一个轻量级的神经网络推理框架,本仓库分离了静态库,使其适合跨平台编译☆21Sep 12, 2024Updated last year
- Notes for EE364a - Convex Optimization I @ Stanford (will update Ch 6 - Ch 13 later)☆12Jul 30, 2019Updated 7 years ago
- Implementation of paper "Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval"☆17Jan 10, 2022Updated 4 years ago
- This is the pytorch implementation of the UAI2023 paper "A Trajectory is Worth Three Sentences: Multimodal Transformer for Offline Reinf…☆11Oct 9, 2023Updated 2 years ago
- Formalizing the Intel 4004 microprocessor☆26Jul 14, 2026Updated last month