This repository is a playground for beginners to learn reinforcement learning. It is a collection of simple environments and agents to get you started with reinforcement learning.
☆26Jul 30, 2024Updated last year
Alternatives and similar repositories for RL_PlayGround
Users that are interested in RL_PlayGround are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Keras 1D Depthwise Convolutional layer☆10May 22, 2020Updated 6 years ago
- Exploration of techniques to solve tasks with a Panda robotic arm. Simulation based on PyBullet physics engine and gymnasium.☆10Mar 17, 2025Updated last year
- ☆11Feb 17, 2025Updated last year
- Time series data structure learning with NOTEARS and DYNOTEARS☆15May 23, 2024Updated 2 years ago
- This set of codes implements our TSG paper "Hierarchical Deep Learning Model for Degradation Prediction per Look-Ahead Scheduled Battery …☆11Feb 24, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆24Oct 27, 2022Updated 3 years ago
- Minimizing Carbon Emission during EV Charging☆15Jul 1, 2022Updated 4 years ago
- Neural Networks package for R with a fast C++ back-end and special support for unsupervised anomaly detection using autoencoders☆12Oct 9, 2025Updated 9 months ago
- A MAPF Algorithm Visualizer☆11Mar 2, 2025Updated last year
- ☆12Mar 15, 2023Updated 3 years ago
- RLC4CLR employs curriculum learning to train a reinforcement learning controller (RLC) for a distribution system critical load restoratio…☆22Jul 3, 2025Updated last year
- Learning how to ride a bicycle using reinforcement learning.☆13Dec 11, 2013Updated 12 years ago
- 智能控制结课作业实验代码实现部分,包括模糊控制器和PID控制器实现以及控制器参数优化整定,PID参数采用Nelder-Mead优化,模糊控制器参数采用遗传算法优化。☆10Dec 2, 2024Updated last year
- [AAMAS 2024] HiMAP: Learning Heuristics-Informed Policies for Large-Scale Multi-Agent Pathfinding☆14Mar 12, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- This repository is the demo implementation of [Deep Dimension Reduction for Supervised Representation Learning].☆12Sep 12, 2024Updated last year
- DiffClass: Diffusion-Based Class Incremental Learning☆12Sep 26, 2024Updated last year
- ☆13Aug 17, 2020Updated 5 years ago
- DC Optimal Power Flow (OPF) via gurobi and yalmip, respectively. An example for learning gurobi and yalmip. 分别通过gurobi和yalmip实现直流最优潮流。学习g…☆12Jul 29, 2025Updated 11 months ago
- 一般印象,flask 项目适合做一些短小精悍的项目,特别是与 sqlite、mysql 等数据库结合很是般配。但是在一些大公司,特别是一些金融行业等国企公司,还是以 oracle 居多,那么,这个小辣椒(flask)就无用武之地了吗?No, No, No... 下面将以 …☆11Jan 26, 2018Updated 8 years ago
- TS_SPMA: The Tabu Search algorithm for simultaneous scheduling problem of machines and AGVs.☆12Apr 30, 2021Updated 5 years ago
- Aqua Metropolis University Stargazer☆20Mar 9, 2026Updated 4 months ago
- Code for "Score-based Generative Modeling Secretly Minimizes the Wasserstein Distance", NeurIPS 2022☆17Feb 11, 2023Updated 3 years ago
- [INFOCOM 2020] Energy-Efficient UAV Crowdsensing with Multiple Charging Stations by Deep Learning☆18May 16, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Robot control with Traditional control and Optimal control and Learning-based control☆10Oct 9, 2022Updated 3 years ago
- This repository contains various Transmission and Generation related optimization problem. The problems are solved using YALMIP toolbox w…☆14Jan 29, 2021Updated 5 years ago
- 复现中国电机工程学报《微电网两阶段鲁棒优化经济调度方法》☆14May 7, 2023Updated 3 years ago
- In this repository, networked control of PI-based controllers for an islanded microgrid has been simulated.☆13Nov 10, 2022Updated 3 years ago
- This repository contains the code and released models for the paper Segmenting Text and Learning Their Rewards for Improved RLHF in Langu…☆19Jan 8, 2025Updated last year
- Application of the Soft Actor-Critic algorithm to the control and guidance of the RexROV 2, based on the UUV Simulator☆16May 27, 2024Updated 2 years ago
- IntelliHealer: An imitation and reinforcement learning platform for self-healing distribution networks☆34Oct 24, 2025Updated 8 months ago
- Mobile Manipulator control package for beverage serving automation☆11Dec 1, 2019Updated 6 years ago
- building electricity load day-ahead forecast (15min interval)☆15Oct 28, 2017Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 基于2016年电工杯数学建模竞赛数据集建立的超短期以及短期负荷预测☆18May 4, 2024Updated 2 years ago
- simulator for agv scheduling system (Project 2023, SJTU)☆14May 22, 2023Updated 3 years ago
- A simulink platform for Overactuated ROV (BlueROV Heav): dynamic positioning, path following...☆18Apr 19, 2024Updated 2 years ago
- cubemx stm32h743iit6, usb device uvc camera☆16Apr 14, 2021Updated 5 years ago
- NILM (Non-intrusive Load Monitoring) 实验代码与结果 本仓库包含使用深度学习模型(LSTM、RNN 和 MultiHead Transformer)进行非侵入式负荷检测的代码和实验结果。NILM 技术用于从总负荷数据中提取各个电器的用电信…☆19Oct 21, 2024Updated last year
- Benchmarking Large Neighborhood Search for Multi-Agent Path Finding☆21Aug 18, 2025Updated 11 months ago
- Robust Orthonormal Subspace Learning in Python☆15Jun 1, 2020Updated 6 years ago