深度强化学习各算法介绍与Pytorch实现
☆77Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for rl-notebook
Users that are interested in rl-notebook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Code for Paper “Relay Hindsight Experience Replay: Self-Guided Continual Reinforcement Learning for Sequential Object Manipulation Ta…☆160Jul 10, 2024Updated 2 years ago
- robopal: a multi-platform, modular robot simulation framework based on MuJoCo, mainly used for reinforcement learning and control algori…☆307May 27, 2025Updated last year
- 基于gym的pytorch深度强化学习(DRL)(PPO,PPG,DQN,SAC,DDPG,TD3等算法)☆155Jan 23, 2026Updated 7 months ago
- RLBench simulation project for autonomous bin picking using Pandas robot arm☆11Mar 1, 2021Updated 5 years ago
- Implementation of Soft Actor-Critic with Hindsight Experience Replay☆21Oct 23, 2020Updated 5 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Deep reinforcement learning with a particle dynamics environment applied to emergency evacuation of a room with obstacles☆10Mar 6, 2026Updated 6 months ago
- 2D Grid Environment with common utils (raytracing) and quadrotor dynamics. With Exponential Control Barrier Functions☆13Jun 3, 2020Updated 6 years ago
- 利用深度强化学习的方法实现多智能体间离散无交流的障碍避免。其中强化学习算法训练模型所需的数据集由最优互惠碰撞避免(Optimal Reciprocal Collision Avoidance, ORCA)算法生成。☆89Mar 14, 2019Updated 7 years ago
- Exploration of techniques to solve tasks with a Panda robotic arm. Simulation based on PyBullet physics engine and gymnasium.☆10Mar 17, 2025Updated last year
- The purpose of this project is to implement machine learning methods to study resource allocation problems, that is how to share limited …☆16Jun 7, 2022Updated 4 years ago
- ☆17Aug 6, 2024Updated 2 years ago
- 本仓库包含了完整的深度学习应用开发流程,以经典的手写字符识别为例,基于LeNet网络构建。推理部分使用torch、onnxruntime以及openvino框架💖☆17Apr 20, 2026Updated 5 months ago
- ☆10Oct 20, 2021Updated 4 years ago
- Code for PolyTask: Learning Unified Policies through Behavior Distillation☆11Oct 13, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- pybullet WBC quadruped robot☆13Apr 5, 2022Updated 4 years ago
- Python 3.x implementation of the retargeting system presented in the paper "Task Oriented Hand Motion Retargeting for Dexterous Manipulat…☆16Jul 10, 2022Updated 4 years ago
- 深度强化学习路径规划, SAC路径规划, Soft Actor-Critic算法, SAC-pytorch,激光雷达Lidar避障,激光雷达仿真,Adaptive-SAC☆598Dec 3, 2025Updated 9 months ago
- A cell counter using computer vision techniques.☆10May 13, 2022Updated 4 years ago
- Mobile Robot Path Planning and Obstacle Avoidance Using PSO in Python☆54Mar 13, 2023Updated 3 years ago
- ☆27Jan 4, 2026Updated 8 months ago
- DRLib:a Concise Deep Reinforcement Learning Library, Integrating HER, PER and D2SR for Almost Off-Policy RL Algorithms.☆561Apr 2, 2024Updated 2 years ago
- Multi-Agent Deep Recurrent Q-Learning with Bayesian epsilon-greedy on AirSim simulator☆13Apr 1, 2022Updated 4 years ago
- ☆12Jan 3, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆20Sep 6, 2024Updated 2 years ago
- The following project is an implementation of "Hakan Ülker, Cemal Baykara, Can Özsoy, "Design of MPCs for a fixed wing UAV", Aircraft Eng…☆12Mar 18, 2022Updated 4 years ago
- 改进遗传算法的钢厂多行车调度☆12Jun 20, 2019Updated 7 years ago
- gym_fetch_env with insert drawer open door☆13Mar 22, 2022Updated 4 years ago
- Virtual RobotX Repository☆14Dec 7, 2019Updated 6 years ago
- A pathfinding application of the GWO heuristic algorithm☆11Feb 4, 2020Updated 6 years ago
- ☆11Feb 17, 2025Updated last year
- This repo refers to paper Invariant Transform Experience Replay. And this repo is built on top of OpenAI Baseline. For more information p…☆12Feb 2, 2021Updated 5 years ago
- The test code for the paper "Attention-based advantage actor-critic algorithm with prioritized experience replay for complex 2-D robotic …☆10Aug 7, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A generic Python interface class for robot simulations using PyBullet. Also provides an IK interface for multi-end-effector robots that u…☆29Sep 11, 2026Updated last week
- environments for reinforcement learning based on panda-gym☆19Aug 22, 2022Updated 4 years ago
- A unified framework of Deep Reinforcement Learning and Deep Imitation Learning in simulation environments☆15Nov 11, 2019Updated 6 years ago
- ☆14Nov 4, 2022Updated 3 years ago
- Open Source Code for RA-L 2025 Paper☆69Oct 1, 2025Updated 11 months ago
- PPO with Hindsight Experience Replay (HER)☆12May 8, 2018Updated 8 years ago
- 预测-校正学习计算制导律☆13Jun 22, 2021Updated 5 years ago