深度强化学习各算法介绍与Pytorch实现
☆77Jul 18, 2024Updated 2 years ago
Alternatives and similar repositories for rl-notebook
Users that are interested in rl-notebook are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The Code for Paper “Relay Hindsight Experience Replay: Self-Guided Continual Reinforcement Learning for Sequential Object Manipulation Ta…☆160Jul 10, 2024Updated 2 years ago
- robopal: a multi-platform, modular robot simulation framework based on MuJoCo, mainly used for reinforcement learning and control algori…☆307May 27, 2025Updated last year
- RLBench simulation project for autonomous bin picking using Pandas robot arm☆11Mar 1, 2021Updated 5 years ago
- Path Planning with Reinforcement Learning algorithms in an unknown environment☆22Jan 29, 2026Updated 8 months ago
- 2D Grid Environment with common utils (raytracing) and quadrotor dynamics. With Exponential Control Barrier Functions☆13Jun 3, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 强化学习算法库,包含了目前主流的强化学习算法(Value based and Policy based)的代码,代码都经过调试并可以运行☆126Nov 2, 2023Updated 2 years ago
- 利用 深度强化学习的方法实现多智能体间离散无交流的障碍避免。其中强化学习算法训练模型所需的数据集由最优互惠碰撞避免(Optimal Reciprocal Collision Avoidance, ORCA)算法生成。☆90Mar 14, 2019Updated 7 years ago
- Exploration of techniques to solve tasks with a Panda robotic arm. Simulation based on PyBullet physics engine and gymnasium.☆10Mar 17, 2025Updated last year
- The purpose of this project is to implement machine learning methods to study resource allocation problems, that is how to share limited …☆16Jun 7, 2022Updated 4 years ago
- RL for path planning☆13Aug 4, 2018Updated 8 years ago
- 基于优化算法的人员应急疏散优化方案 | Optimization Plan for Emergency Evacuation of Personnel Based on Optimization Algorithm☆13Sep 4, 2024Updated 2 years ago
- 本仓库包含了完整的深度学习应用开发流程,以经典的手写字符识别为例,基于LeNet网络构建。推理部分使用torch、onnxruntime以及openvino框架💖☆17Apr 20, 2026Updated 5 months ago
- ☆11Oct 20, 2021Updated 4 years ago
- Code for PolyTask: Learning Unified Policies through Behavior Distillation☆11Oct 13, 2023Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- pybullet WBC quadruped robot☆13Apr 5, 2022Updated 4 years ago
- Python 3.x implementation of the retargeting system presented in the paper "Task Oriented Hand Motion Retargeting for Dexterous Manipulat…☆16Jul 10, 2022Updated 4 years ago
- 深度强化学习路径规划, SAC路径规划, Soft Actor-Critic算法, SAC-pytorch,激光雷达Lidar避障,激光雷达仿真,Adaptive-SAC☆602Dec 3, 2025Updated 10 months ago
- [ACMMM 2024] Consistent123: One Image to Highly Consistent 3D Asset Using Case-Aware Diffusion Priors☆25Oct 22, 2024Updated last year
- code for the paper Offline Prioritized Experience Replay☆12Jun 13, 2023Updated 3 years ago
- A cell counter using computer vision techniques.☆10May 13, 2022Updated 4 years ago
- ☆27Jan 4, 2026Updated 9 months ago
- DRLib:a Concise Deep Reinforcement Learning Library, Integrating HER, PER and D2SR for Almost Off-Policy RL Algorithms.☆561Apr 2, 2024Updated 2 years ago
- Multi-Agent Deep Recurrent Q-Learning with Bayesian epsilon-greedy on AirSim simulator☆13Apr 1, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆12Jan 3, 2020Updated 6 years ago
- ☆20Sep 6, 2024Updated 2 years ago
- The following project is an implementation of "Hakan Ülker, Cemal Baykara, Can Özsoy, "Design of MPCs for a fixed wing UAV", Aircraft Eng…☆12Mar 18, 2022Updated 4 years ago
- 改进遗传算法的钢厂多行车调度☆12Jun 20, 2019Updated 7 years ago
- Deep Recurrent Q-Network with different exploration strategies for self-driving cars (using AirSim)☆10Sep 5, 2024Updated 2 years ago
- gym_fetch_env with insert drawer open door☆13Mar 22, 2022Updated 4 years ago
- Source code of "Variational Imitation Learning with Diverse-quality Demonstrations" in ICML 2020. This github repository includes python …☆20Aug 16, 2021Updated 5 years ago
- Virtual RobotX Repository☆14Dec 7, 2019Updated 6 years ago
- ☆11Feb 17, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repo refers to paper Invariant Transform Experience Replay. And this repo is built on top of OpenAI Baseline. For more information p…☆12Feb 2, 2021Updated 5 years ago
- Matlab codes for paper entitled "Distributed Hybrid Consensus-Based Square-Root Cubature Quadrature Information Filter and Its Applicatio…☆12Oct 31, 2019Updated 6 years ago
- environments for reinforcement learning based on panda-gym☆19Aug 22, 2022Updated 4 years ago
- A unified framework of Deep Reinforcement Learning and Deep Imitation Learning in simulation environments☆15Nov 11, 2019Updated 6 years ago
- ☆14Nov 4, 2022Updated 3 years ago
- Codes for "Quantitative Comparison of Reinforcement Learning and Data-driven Model Predictive Control for Chemical and Biological Process…☆12Dec 18, 2023Updated 2 years ago
- Open Source Code for RA-L 2025 Paper☆69Oct 1, 2025Updated last year