《深度强化学习:原理与实践》,Code of the book <Deep Reinforcement Learning: Principles and Practices>
☆211Apr 4, 2019Updated 7 years ago
Alternatives and similar repositories for Deep-Reinforcement-Learning
Users that are interested in Deep-Reinforcement-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 《深度学习原理与实践》相关代码——source code of the book <deep learning in action>☆68Aug 24, 2018Updated 7 years ago
- MATLAB implementation of DQN for a navigation environment☆13Aug 13, 2020Updated 6 years ago
- ☆329Jul 1, 2025Updated last year
- ☆12Jun 22, 2023Updated 3 years ago
- AI Infra主要是指AI的基础建设,包括AI芯片、AI编译器、AI推理和训练框架等AI全栈底层技术。☆270Mar 26, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Improving coordinated (two intersections) transit signal priority on bus travel time and headway reliability with single agent reinforcem…☆14Oct 2, 2021Updated 4 years ago
- 用深度学习+强化学习编写的一个五子棋人工智障☆45Feb 16, 2018Updated 8 years ago
- Written by Dr Trang Mai (maicongtrang@gmail.com)☆26Jun 3, 2020Updated 6 years ago
- 强化学习中文教程(蘑菇书🍄),在线阅读地址:https://datawhalechina.github.io/easy-rl/☆14,540Dec 30, 2025Updated 7 months ago
- ☆32May 29, 2019Updated 7 years ago
- 这是一个学习强化学习基础原理的仓库,主要包括了《深入浅出强化学习原理入门》书中一些例子和课后作业的代码☆272Dec 4, 2018Updated 7 years ago
- [JSS'19] A Blockchain-based Lightweight Framework for Edge and Fog Computing☆45Jun 18, 2021Updated 5 years ago
- Reference implementations of (my) control algorithms for Markov Jump Linear Systems without mode observation.☆14Sep 15, 2016Updated 9 years ago
- 基于深度强化学习的资源调度研究☆92Feb 9, 2018Updated 8 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Intro to Reinforcement Learning (强化学习纲要)☆3,597Jul 25, 2020Updated 6 years ago
- ☆18Dec 5, 2017Updated 8 years ago
- Temporal Pattern Attention for Multivariate Time Series Forecasting☆16Jan 27, 2021Updated 5 years ago
- 中文整理的强化学习资料(Reinforcement Learning)☆2,189Apr 30, 2020Updated 6 years ago
- LeetCode collection☆43Nov 25, 2020Updated 5 years ago
- An implementation of Deep Q-Learning from Demonstrations (DQfD) for playing Atari 2600 video games☆31Dec 10, 2022Updated 3 years ago
- This project provides a set of translators to convert OpenAI Gym environments into text-based environments. It is designed to investigate…☆22May 29, 2024Updated 2 years ago
- This repository contains code related to solving and visualizing the Bi-Directional Electric Vehicle Routing Problem (B-EVRP) as well as …☆17Apr 1, 2022Updated 4 years ago
- ☆17Mar 15, 2021Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- RandomWarpingSeries (RWS) is a simple code for generating the vector representation of time-series for time-series classification, cluste…☆21Jan 21, 2019Updated 7 years ago
- ROS1 node for the Blueprint Oculus front scan sonar.☆13Jan 4, 2023Updated 3 years ago
- Guided policy search in Python and ROS Indigo.☆26Feb 12, 2026Updated 6 months ago
- Convolutional Spiking Neural Network to recognize speech utterances using Spike-Timing-Dependent Plasticity☆10Mar 9, 2021Updated 5 years ago
- Simple Reinforcement learning tutorials, 莫烦Python 中文AI教学☆9,504Mar 31, 2024Updated 2 years ago
- You will learn how to remove periodic noise in the Fourier domain☆12Oct 17, 2018Updated 7 years ago
- Non-convex optimization using proximal methods☆14Mar 22, 2018Updated 8 years ago
- 2022年华为软件精英赛初赛☆11Apr 2, 2022Updated 4 years ago
- PipelineLLM 是一个系统性的大语言模型(LLM)后训练学习项目,涵盖从监督微调(SFT)到偏好优化(DPO)、强化学习(RLHF/PPO/GRPO)再到持续学习(Continual Learning)的完整技术栈。☆33Jan 16, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Tiny Calculator with support of +, -, *, /, ^, sin, cos, tan...☆10Apr 2, 2024Updated 2 years ago
- 深度学习、强化学习、模仿学习与机器人☆482Oct 31, 2020Updated 5 years ago
- notes for NJU courses☆18Oct 26, 2021Updated 4 years ago
- Homomorphic encryption library for encrypted control☆13Aug 9, 2024Updated 2 years ago
- ☆12Jan 31, 2022Updated 4 years ago
- Keras Implementation of TD3(Twin Delayed DDPG) with PER(Prioritized Experience Replay) option on OpenAI gym framework☆11May 29, 2021Updated 5 years ago
- WISHBONE Interconnect☆11Oct 1, 2017Updated 8 years ago