Reinforcement Learning algorithms and use-cases, including DQN, PG, A3C, PPO etc. and RLHF, AlphaZero implementations. Designed for clarity, ease of use, and educational purposes.
☆55May 29, 2024Updated 2 years ago
Alternatives and similar repositories for CleanRL
Users that are interested in CleanRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- awesome-edge-computing,边缘计算各种资料汇总,相关技术资料汇总☆23Nov 8, 2021Updated 4 years ago
- On-Policy Policy Gradient Algorithms in JAX☆44Jan 25, 2024Updated 2 years ago
- Code for Policy Bifurcation in Safe Reinforcement Learning☆10Jul 4, 2025Updated last year
- Deep Q Network for Multi-agent RL☆15Oct 18, 2020Updated 5 years ago
- [CVPR2025] Hand-held Object Reconstruction from RGB Video with Dynamic Interaction☆37Sep 1, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- NeurIPS'23: Energy Discrepancies: A Score-Independent Loss for Energy-Based Models☆18Oct 22, 2024Updated last year
- Benchmark for Multi-robot Cleaning Task Allocation☆14Aug 13, 2023Updated 3 years ago
- Markov Chain Monte Carlo (MCMC) and importance sampling in the context of Bayesian linear regression☆11Feb 25, 2018Updated 8 years ago
- Build a bridge that connects beginners to deep reinforcement learning.☆10Sep 23, 2024Updated last year
- db-LaCAM: Fast and Scalable Multi-Robot Kinodynamic Motion Planning with Discontinuity-Bounded Search and Lightweight MAPF (ICAPS-26)☆27Updated this week
- MIGSAA Project 2 - Langevin Monte Carlo Algorithms☆15Jul 25, 2023Updated 3 years ago
- NeurIPS2022: Constrained Update Projection Approach to Safe Policy Optimization☆13Apr 10, 2023Updated 3 years ago
- Thesis in Federated Learning using an Edge/Cloud Computing architecture☆10Feb 26, 2021Updated 5 years ago
- ☆20Jan 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of the paper "Guidance Graph Optimization for Lifelong Multi-Agent Path Finding", published in IJCAI 2024.☆26Mar 10, 2026Updated 6 months ago
- A SITL guide for setting up Ardupilot, Gazebo & ROS☆16Jul 27, 2020Updated 6 years ago
- [CoRL 2022] Official implementation of the publication Residual Skill Policies: Learning an Adaptable Skill-based Action Space for Reinfo…☆26Jan 3, 2023Updated 3 years ago
- Policy Transfer across Visual and Dynamics Domain Gaps via Iterative Grounding (RSS 2021)☆12Oct 22, 2021Updated 4 years ago
- The model for edge classification by transforming edges to nodes.☆15Dec 22, 2020Updated 5 years ago
- A system for running Multi-Agent Path Finding (MAPF) experiments, with multiple implemented algorithms.☆39Apr 18, 2026Updated 4 months ago
- [ACM MM'25] Code for the paper "Open3D-VQA: A Benchmark for Embodied Spatial Reasoning with Multimodal Large Language Model in Open Space…☆18Jul 9, 2026Updated 2 months ago
- A Monte-Carlo simulator for Mobile Edge/Cloud Computing☆12Aug 22, 2023Updated 3 years ago
- Offline RLHF codebase implementation for "Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human …☆42Mar 26, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The CLI & python API for the well-known project gpt-academic.☆19Sep 22, 2024Updated last year
- ☆26Jan 20, 2022Updated 4 years ago
- [CVPR 2025 Highlight] CAP-Net: A Unified Network for 6D Pose and Size Estimation of Categorical Articulated Parts from a Single RGB-D Ima…☆46Aug 19, 2025Updated last year
- Distributed Uplink Beamforming in Cell-Free Networks Using Deep Reinforcement Learning☆10Mar 20, 2021Updated 5 years ago
- Dubin's Vehicle Model in Gym Environment for Path Tracking using RL Algorithms☆16Jun 30, 2021Updated 5 years ago
- RL and MARL from Mobile Edge Computing Load Optimization☆12Jun 28, 2023Updated 3 years ago
- Dynamic Attention Encoder-Decoder model to learn and design heuristics to solve capacitated vehicle routing problems☆50Jan 7, 2021Updated 5 years ago
- ☆45Jul 24, 2024Updated 2 years ago
- This repository contains a Reinforcement learning algorithm for task scheduling in edge-cloud computing.☆13Feb 10, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- On-Policy Model-free Reinforcement Learning for simplified Blackjack (David Silver Assignement)☆11Nov 20, 2017Updated 8 years ago
- The code for the paper "A Bayesian Approach to Online Planning" published in ICML 2024.☆13Jun 17, 2024Updated 2 years ago
- Simulation Design of a Robotic Mobile Manipulator with Drone in Isaacsim.☆14Oct 8, 2024Updated last year
- [CVPR 2026] Few-Shot Incremental 3D Object Detection in Dynamic Indoor Environments☆18Aug 24, 2026Updated 3 weeks ago
- ☆13May 4, 2023Updated 3 years ago
- LVI-SAM for easier using (更简单的使用LVI-SAM的方法)☆11Dec 5, 2023Updated 2 years ago
- Simulator for evaluating cloud/edge requests from connected vehicles and computing statistical analysis of the input network☆12May 7, 2018Updated 8 years ago