深度强化学习各算法介绍与Pytorch实现
☆77Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for rl-notebook
Users that are interested in rl-notebook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Code for Paper “Relay Hindsight Experience Replay: Self-Guided Continual Reinforcement Learning for Sequential Object Manipulation Ta…☆159Jul 10, 2024Updated 2 years ago
- robopal: a multi-platform, modular robot simulation framework based on MuJoCo, mainly used for reinforcement learning and control algori…☆304May 27, 2025Updated last year
- 基于gym的pytorch深度强化学习(DRL)(PPO,PPG,DQN,SAC,DDPG,TD3等算法)☆153Jan 23, 2026Updated 6 months ago
- RLBench simulation project for autonomous bin picking using Pandas robot arm☆11Mar 1, 2021Updated 5 years ago
- 2D Grid Environment with common utils (raytracing) and quadrotor dynamics. With Exponential Control Barrier Functions☆13Jun 3, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 利用深度强化学习的方法实现多智能体间离散无交流的障碍避免。其中强化学习算法训练模型所需的数据集由最优互惠碰撞避免(Optimal Reciprocal Collision Avoidance, ORCA)算法生成。☆89Mar 14, 2019Updated 7 years ago
- Control inverted pendulum by LQR in OpenAI Gym☆12Oct 2, 2024Updated last year
- Exploration of techniques to solve tasks with a Panda robotic arm. Simulation based on PyBullet physics engine and gymnasium.☆10Mar 17, 2025Updated last year
- ☆17Aug 6, 2024Updated last year
- 本仓库包含了完整的深度学习应用开发流程,以经典的手写 字符识别为例,基于LeNet网络构建。推理部分使用torch、onnxruntime以及openvino框架💖☆17Apr 20, 2026Updated 3 months ago
- ☆10Oct 20, 2021Updated 4 years ago
- Code for PolyTask: Learning Unified Policies through Behavior Distillation☆11Oct 13, 2023Updated 2 years ago
- pybullet WBC quadruped robot☆13Apr 5, 2022Updated 4 years ago
- 深度强化学习路径规划, SAC路径规划, Soft Actor-Critic算法, SAC-pytorch,激光雷达Lidar避障,激光雷达仿真,Adaptive-SAC☆583Dec 3, 2025Updated 7 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- MFO 3D path planning☆15Aug 14, 2019Updated 6 years ago
- code for the paper Offline Prioritized Experience Replay☆12Jun 13, 2023Updated 3 years ago
- A cell counter using computer vision techniques.☆10May 13, 2022Updated 4 years ago
- Mobile Robot Path Planning and Obstacle Avoidance Using PSO in Python☆54Mar 13, 2023Updated 3 years ago
- ☆25Jan 4, 2026Updated 6 months ago
- DRLib:a Concise Deep Reinforcement Learning Library, Integrating HER, PER and D2SR for Almost Off-Policy RL Algorithms.☆564Apr 2, 2024Updated 2 years ago
- Multi-Agent Deep Recurrent Q-Learning with Bayesian epsilon-greedy on AirSim simulator☆13Apr 1, 2022Updated 4 years ago
- ☆12Jan 3, 2020Updated 6 years ago
- ☆19Sep 6, 2024Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 改进遗传算法的钢厂多行车调度☆12Jun 20, 2019Updated 7 years ago
- A set of Matlab/Octave files that performs a method of Nonlinear System Identification.☆26Oct 26, 2018Updated 7 years ago
- Deep Recurrent Q-Network with different exploration strategies for self-driving cars (using AirSim)☆10Sep 5, 2024Updated last year
- gym_fetch_env with insert drawer open door☆13Mar 22, 2022Updated 4 years ago
- ☆19Jan 8, 2026Updated 6 months ago
- Virtual RobotX Repository☆13Dec 7, 2019Updated 6 years ago
- A pathfinding application of the GWO heuristic algorithm☆11Feb 4, 2020Updated 6 years ago
- ☆11Feb 17, 2025Updated last year
- This repo refers to paper Invariant Transform Experience Replay. And this repo is built on top of OpenAI Baseline. For more information p…☆12Feb 2, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The test code for the paper "Attention-based advantage actor-critic algorithm with prioritized experience replay for complex 2-D robotic …☆10Aug 7, 2022Updated 3 years ago
- Matlab codes for paper entitled "Distributed Hybrid Consensus-Based Square-Root Cubature Quadrature Information Filter and Its Applicatio…☆12Oct 31, 2019Updated 6 years ago
- environments for reinforcement learning based on panda-gym☆19Aug 22, 2022Updated 3 years ago
- Open Source Code for RA-L 2025 Paper☆68Oct 1, 2025Updated 9 months ago
- ☆10Nov 16, 2023Updated 2 years ago
- ☆14Nov 4, 2022Updated 3 years ago
- PPO with Hindsight Experience Replay (HER)☆12May 8, 2018Updated 8 years ago