Repository of notes, code and notebooks in Python for the book "Reinforcement Learning: An Introduction" by Richard S. Sutton and Andrew G. Barto
☆40Mar 6, 2026Updated 5 months ago
Alternatives and similar repositories for reinforcement-learning
Users that are interested in reinforcement-learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repository for the 2024 Telluride Topic Area Neuromorphic Systems for Space☆13Jun 21, 2024Updated 2 years ago
- 用koa、c++写的量化交易系统☆14Oct 9, 2017Updated 8 years ago
- Repository of notes, code and notebooks in Python for the book Pattern Recognition and Machine Learning by Christopher Bishop☆2,625Jul 25, 2022Updated 4 years ago
- We open-source our layout level fast EM simulation tool, EMSim, to the public.☆15Feb 8, 2024Updated 2 years ago
- Create Custom GYM Environment for SUMO and reinforcement learning agant☆15May 5, 2023Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Implementation of Proximal Policy Optimization (PPO) for continuous action space (`Pendulum-v1` from gym) using tensorflow2.x and pytorch…☆12Aug 8, 2022Updated 4 years ago
- Grokking on modular arithmetic in less than 150 epochs in MLX☆15Oct 24, 2024Updated last year
- FMCW LiDAR implementation in CARLA simulator☆19Mar 18, 2024Updated 2 years ago
- Exploring HMM, LSTM and Regression techniques to predict respiratory rate of an individual from accelerometer data.☆15Dec 4, 2018Updated 7 years ago
- 用Java实现一个完整的TCP/IP协议,包括一些常见的 应用(ARP, DNS, ICMP, traceroute, ping, IP, UDP,TCP, HTTP,DHCP, TFTP, FTP )☆15Feb 5, 2022Updated 4 years ago
- Providing the answer to "How to do patching on all available SAEs on GPT-2?". It is an official repository of the implementation of the p…☆14Jan 26, 2025Updated last year
- This is the implementation of CounterCurate, the data curation pipeline of both physical and semantic counterfactual image-caption pairs.☆19Jun 27, 2024Updated 2 years ago
- A Deep-Reinforcement-Learning-Based Scheduler for FPGA HLS☆15Feb 27, 2021Updated 5 years ago
- Collated optimization models from numerous sources☆13Jun 18, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Recent papers on Graph Neural Networks-based Recommender System.☆12Aug 21, 2023Updated 2 years ago
- ☆12Mar 1, 2022Updated 4 years ago
- Adaptive Machine Learning-Based Stock Prediction using Financial Time Series Technical Indicators☆10Dec 21, 2019Updated 6 years ago
- My Python Intel 4004 Emulator☆19Jan 29, 2016Updated 10 years ago
- Deep and online learning with spiking neural networks in Python for Graphcore IPU☆10Apr 8, 2023Updated 3 years ago
- A simple tutorial to add medical reasoning using GRPO☆21Feb 10, 2025Updated last year
- ☆11Apr 28, 2024Updated 2 years ago
- Summer School Week 1 & 2 repo☆12Jul 1, 2022Updated 4 years ago
- Landing repository for the paper "Softpick: No Attention Sink, No Massive Activations with Rectified Softmax"☆92Sep 12, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Deep Reinforcement Learning model for high volume and frequency Forex Portfolio Management☆13Jan 11, 2023Updated 3 years ago
- A framework for creating your own reinforcement learning environments using pybullet☆21Oct 7, 2019Updated 6 years ago
- ☆16Nov 14, 2025Updated 8 months ago
- ☆23Nov 27, 2024Updated last year
- gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI☆17Jan 12, 2026Updated 7 months ago
- The ecosystem of geospatial machine learning tools in the Pangeo world.☆12Mar 17, 2025Updated last year
- Submission template for Tiny Tapeout 04☆17Jun 15, 2024Updated 2 years ago
- [TNNLS 2024] Implementation of "TCJA-SNN: Temporal-Channel Joint Attention for Spiking Neural Networks"☆62Apr 16, 2024Updated 2 years ago
- Implementation of paper "Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval"☆17Jan 10, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆24Aug 26, 2017Updated 8 years ago
- ☆16May 14, 2025Updated last year
- Teaching materials for BayesCog workshop, UKE Hamburg (Part 1).☆15Dec 4, 2023Updated 2 years ago
- Formalizing the Intel 4004 microprocessor☆25Jul 14, 2026Updated 3 weeks ago
- [ICLR 2025] Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning (SASR)☆12Aug 26, 2025Updated 11 months ago
- Code for our paper "Decomposing The Dark Matter of Sparse Autoencoders"☆23Feb 6, 2025Updated last year
- Explainability of Deep RL algorithms using graph networks and layer-wise relevance propagation.☆12Aug 20, 2024Updated last year