深度强化学习各算法介绍与Pytorch实现
☆77Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for rl-notebook
Users that are interested in rl-notebook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Code for Paper “Relay Hindsight Experience Replay: Self-Guided Continual Reinforcement Learning for Sequential Object Manipulation Ta…☆160Jul 10, 2024Updated 2 years ago
- robopal: a multi-platform, modular robot simulation framework based on MuJoCo, mainly used for reinforcement learning and control algori…☆304May 27, 2025Updated last year
- 基于gym的pytorch深度强化学习(DRL)(PPO,PPG,DQN,SAC,DDPG,TD3等算法)☆155Jan 23, 2026Updated 7 months ago
- RLBench simulation project for autonomous bin picking using Pandas robot arm☆11Mar 1, 2021Updated 5 years ago
- Implementation of Soft Actor-Critic with Hindsight Experience Replay☆21Oct 23, 2020Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 2D Grid Environment with common utils (raytracing) and quadrotor dynamics. With Exponential Control Barrier Functions☆13Jun 3, 2020Updated 6 years ago
- 利用深度强化学习的方法实现多智能体间离散无交流的障碍避免。其中强化学习算法训练模型所需的数据集由最优互惠碰撞避免(Optimal Reciprocal Collision Avoidance, ORCA)算法生成。☆89Mar 14, 2019Updated 7 years ago
- Exploration of techniques to solve tasks with a Panda robotic arm. Simulation based on PyBullet physics engine and gymnasium.☆10Mar 17, 2025Updated last year
- ☆17Aug 6, 2024Updated 2 years ago
- 本仓库包含了完整的深度学习应用开发流程,以经典的手写字符识别为例,基于LeNet网络构建。推理部分使用torch、onnxruntime以及openvino框架💖☆17Apr 20, 2026Updated 4 months ago
- ☆10Oct 20, 2021Updated 4 years ago
- Code for PolyTask: Learning Unified Policies through Behavior Distillation☆11Oct 13, 2023Updated 2 years ago
- pybullet WBC quadruped robot☆13Apr 5, 2022Updated 4 years ago
- Python 3.x implementation of the retargeting system presented in the paper "Task Oriented Hand Motion Retargeting for Dexterous Manipulat…☆16Jul 10, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 深度强化学习路径规划, SAC路径规划, Soft Actor-Critic算法, SAC-pytorch,激光雷达Lidar避障,激光雷达仿真,Adaptive-SAC☆591Dec 3, 2025Updated 8 months ago
- a flexibility oriented stochastic scheduling framework is presented to evaluate short-term reliability and economic of islanded microgri…☆12Apr 29, 2022Updated 4 years ago
- A cell counter using computer vision techniques.☆10May 13, 2022Updated 4 years ago
- ☆25Jan 4, 2026Updated 7 months ago
- A Neural Network Architecture for the Analysis of Unlabeled Time-Series Data☆10Jun 25, 2019Updated 7 years ago
- DRLib:a Concise Deep Reinforcement Learning Library, Integrating HER, PER and D2SR for Almost Off-Policy RL Algorithms.☆563Apr 2, 2024Updated 2 years ago
- Multi-Agent Deep Recurrent Q-Learning with Bayesian epsilon-greedy on AirSim simulator☆13Apr 1, 2022Updated 4 years ago
- ☆12Jan 3, 2020Updated 6 years ago
- ☆19Sep 6, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Reinforcement Learning Robot avoiding obstacles(Python + V_rep)☆12Oct 29, 2019Updated 6 years ago
- 改进遗传算法的钢厂多行车调度☆12Jun 20, 2019Updated 7 years ago
- Deep Recurrent Q-Network with different exploration strategies for self-driving cars (using AirSim)☆10Sep 5, 2024Updated last year
- gym_fetch_env with insert drawer open door☆13Mar 22, 2022Updated 4 years ago
- Python Implementation of Reinforcement Learning: An Introduction☆30Sep 12, 2019Updated 6 years ago
- ☆20Jan 8, 2026Updated 7 months ago
- Source code of "Variational Imitation Learning with Diverse-quality Demonstrations" in ICML 2020. This github repository includes python …☆20Aug 16, 2021Updated 5 years ago
- Virtual RobotX Repository☆13Dec 7, 2019Updated 6 years ago
- A pathfinding application of the GWO heuristic algorithm☆11Feb 4, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆11Feb 17, 2025Updated last year
- This repo refers to paper Invariant Transform Experience Replay. And this repo is built on top of OpenAI Baseline. For more information p…☆12Feb 2, 2021Updated 5 years ago
- MoE model with onnx runtime☆62May 5, 2024Updated 2 years ago
- A generic Python interface class for robot simulations using PyBullet. Also provides an IK interface for multi-end-effector robots that u…☆29Aug 5, 2026Updated 3 weeks ago
- code of the paper "Reliability modeling and statistical analysis of accelerated degradation process with memory effects and unit-to-unit …☆13Jan 28, 2026Updated 7 months ago
- environments for reinforcement learning based on panda-gym☆19Aug 22, 2022Updated 4 years ago
- PyTorch implementation of Munchausen Reinforcement Learning based on DQN and SAC. Handles discrete and continuous action spaces☆15Oct 3, 2021Updated 4 years ago