Python Implementation of Reinforcement Learning: An Introduction
☆30Sep 12, 2019Updated 6 years ago
Alternatives and similar repositories for reinforcement-learning-an-introduction
Users that are interested in reinforcement-learning-an-introduction are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Anti exploration in offline reinforcement learning☆11May 17, 2021Updated 5 years ago
- BasicRL: easy and fundamental codes for deep reinforcement learning。It is an improvement on rainbow-is-all-you-need and OpenAI Spinning U…☆15Oct 29, 2021Updated 4 years ago
- ☆11Sep 29, 2021Updated 4 years ago
- PyTorch implementation of DQN, AC, ACER, A2C, A3C, PG, DDPG, TRPO, PPO, SAC, TD3 and ....☆53Mar 18, 2020Updated 6 years ago
- ☆23Nov 3, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- An environment based on JSBSIM aimed at one-to-one close air combat.☆14May 15, 2023Updated 3 years ago
- Official codebase for "The Generalization Gap in Offline Reinforcement Learning" accepted to ICLR 2024☆29Apr 8, 2026Updated 3 months ago
- This is a ROS repository to track an underwater target using a Particle Filter range-only method and the SparusII AUV☆11Nov 27, 2024Updated last year
- The implementation of Scalable Safe Multi-Agent Reinforcement Learning for Multi-Agent System.☆11Sep 8, 2025Updated 10 months ago
- ☆11Jan 14, 2026Updated 6 months ago
- ☆34Jun 21, 2024Updated 2 years ago
- Reinforcement Learning Robot avoiding obstacles(Python + V_rep)☆12Oct 29, 2019Updated 6 years ago
- ☆15Jul 4, 2022Updated 4 years ago
- ☆16Dec 15, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Assignments for CS294-112.☆30Sep 11, 2019Updated 6 years ago
- using information theory to encourage agents to cooperate and compete☆19Oct 4, 2018Updated 7 years ago
- The official implementation of "When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning" (ICLR2023)☆44Mar 6, 2023Updated 3 years ago
- Implementation of IMvGCN in our paper: Interpretable Graph Convolutional Network for Multi-view Semi-supervised Learning, IEEE TMM.☆10Mar 25, 2024Updated 2 years ago
- References for factor model☆44Oct 12, 2020Updated 5 years ago
- Official code for ICLR 2024 paper, SEABO: A Simple Search-Based Method for Offline Imitation Learning☆12Jan 19, 2024Updated 2 years ago
- unofficial implementation of https://arxiv.org/pdf/2301.08871v1.pdf on pytorch☆14Apr 20, 2023Updated 3 years ago
- ☆17Aug 6, 2024Updated last year
- ☆65Jan 30, 2026Updated 5 months ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Code for running RL experiments on continuing (non-episodic) problems.☆22Feb 13, 2026Updated 5 months ago
- papers about reinforcement learning☆13Jan 4, 2021Updated 5 years ago
- This repository accompanies the following paper: A Workflow for Offline Model-Free Robotic RL☆13Nov 5, 2021Updated 4 years ago
- code for the paper Offline Prioritized Experience Replay☆12Jun 13, 2023Updated 3 years ago
- Code repository accompanying the Heuristic Guided RL NeurIPS'21 paper☆17Jan 3, 2022Updated 4 years ago
- Facebear's minimal implementation of SBAC (Soft behavior regularized actor critic, NIPS22 offline RL workshop)☆11Jul 4, 2022Updated 4 years ago
- common road TUM竞赛☆13Sep 3, 2021Updated 4 years ago
- ☆11Oct 3, 2022Updated 3 years ago
- ☆12Sep 8, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, T…☆14Feb 2, 2025Updated last year
- Conservative Q learning in Jax☆58Feb 7, 2023Updated 3 years ago
- [KDD 2024] Revisiting Modularity Maximization for Graph Clustering: A Contrastive Learning Perspective☆20Jun 26, 2024Updated 2 years ago
- [IEEE TPAMI] A Framework for Constrained Multi-Objective Reinforcement Learning☆19Apr 18, 2025Updated last year
- Code for paper on ICRA 2022 workshop on Deformable Object Manipulation. In this work we learn keypoints from synthetic data for robotic c…☆15Aug 6, 2024Updated last year
- Compress numerical data using variable byte encoding.☆13Jul 24, 2017Updated 8 years ago
- [PNAS'18] Recurrent computations for visual pattern completion: Classification of occluded images in humans and recurrent neural networks☆19Sep 11, 2018Updated 7 years ago