Chinese Translation for Book 《Reinforcement Learning- An Introduction》-Second Edition
☆127Apr 15, 2019Updated 7 years ago
Alternatives and similar repositories for rl-intro-book-chinese
Users that are interested in rl-intro-book-chinese are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A translation of Reinforcement Learning: An Introduction☆114Aug 20, 2018Updated 8 years ago
- ☆15Aug 24, 2019Updated 7 years ago
- sutton 的增强学习导论中文版翻译☆29Feb 9, 2018Updated 8 years ago
- 《Reinforcement Learning: An Introduction》(第二版)中文翻译☆690Apr 9, 2022Updated 4 years ago
- A Q & A system based on Chinese wikipedia knowledge☆19May 26, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A python implemenation of tabular MuZero for educational purposes☆21Dec 11, 2019Updated 6 years ago
- Implementation of SNAIL(A Simple Neural Attentive Meta-Learner) with Gluon☆12Feb 22, 2019Updated 7 years ago
- Generating NEW Reuters articles from Reuters articles.☆16Jan 10, 2017Updated 9 years ago
- Python Implementation of Reinforcement Learning: An Introduction☆14,758Aug 9, 2024Updated 2 years ago
- discrete gate sizing☆14Nov 23, 2020Updated 5 years ago
- Code for the ICML 2013 and AAAI 2013 papers on ELLA☆13Mar 10, 2017Updated 9 years ago
- ROS package for robot learning☆17Oct 16, 2019Updated 6 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- 中文整理的强化学习资料(Reinforcement Learning)☆2,193Apr 30, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [KDD 2026] Official implementation of "FaST: Efficient and Effective Long-Horizon Forecasting for Large-Scale Spatial-Temporal Graphs via…☆18Jun 1, 2026Updated 2 months ago
- Sparse Convex Optimization Toolkit (SCOT)☆13Feb 5, 2024Updated 2 years ago
- A Spiking Multi-Layer Perceptron☆33Sep 5, 2017Updated 8 years ago
- This project uses Pepper robot for human tracking, RGBD data acquisition and physical interaction with people.☆14Feb 8, 2017Updated 9 years ago
- machine learning trading system using random decision tree to train the technical indicators☆10Apr 11, 2017Updated 9 years ago
- The architecture used to train the level generator in the game Relay.☆12Apr 8, 2017Updated 9 years ago
- 这是一个学习强化学习基础原理的仓库,主要包括了《深入浅出强化学习原理入门》书中一些例子和课后作业的代码☆272Dec 4, 2018Updated 7 years ago
- Multimodal Deep Q-Network (MDQN) for modelling human-like social intelligence.☆14Feb 23, 2017Updated 9 years ago
- Reinforcement Learning, Tutorials in Chinese☆11Jun 9, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repo is using imitating learning to optimize portfolio. The code was derived from https://github.com/vermouth1992/drl-portfolio-mana…☆11May 16, 2019Updated 7 years ago
- ☆20Mar 1, 2019Updated 7 years ago
- A TensorFlow implementation of DeepMind's A Distributional Perspective on Reinforcement Learning.(C51-DQN)☆57Aug 25, 2017Updated 9 years ago
- Matlab Implementation of Autonomous Cross-Domain Knowledge Transfer in Lifelong Policy Gradient Reinforcement Learning☆17Mar 14, 2017Updated 9 years ago
- This is a simple experiment designed to uncover which technical indicators are the most important.☆13Jan 11, 2019Updated 7 years ago
- ☆14Jul 24, 2018Updated 8 years ago
- Visual Navigation with Spatial Attention☆37Jan 2, 2025Updated last year
- Implement BinaryNet of CNN with chainer☆11May 5, 2016Updated 10 years ago
- ☆43Apr 28, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- IJCAI 2019 - Regularized Opponent Model with Maximum Entropy Objective (ROMMEO)☆23Dec 8, 2022Updated 3 years ago
- Implementation of Deepmind's Neural Episodic Control☆59May 9, 2018Updated 8 years ago
- VAEs on Sparse Data☆12Oct 15, 2017Updated 8 years ago
- Intepretability method to find what navigation agents learn☆19Jun 16, 2022Updated 4 years ago
- Resources for Auxiliary Tasks and Exploration Enable ObjectNav☆42Oct 22, 2021Updated 4 years ago
- ☆11Mar 29, 2019Updated 7 years ago
- ☆10Sep 20, 2018Updated 7 years ago