Chinese Translation for Book 《Reinforcement Learning- An Introduction》-Second Edition
☆127Apr 15, 2019Updated 7 years ago
Alternatives and similar repositories for rl-intro-book-chinese
Users that are interested in rl-intro-book-chinese are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A translation of Reinforcement Learning: An Introduction☆114Aug 20, 2018Updated 7 years ago
- ☆15Aug 24, 2019Updated 6 years ago
- sutton 的增强学习导论中文版翻译☆29Feb 9, 2018Updated 8 years ago
- 《Reinforcement Learning: An Introduction》(第二版)中文翻译☆686Apr 9, 2022Updated 4 years ago
- A Q & A system based on Chinese wikipedia knowledge☆19May 26, 2017Updated 9 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A python implemenation of tabular MuZero for educational purposes☆21Dec 11, 2019Updated 6 years ago
- Implementation of SNAIL(A Simple Neural Attentive Meta-Learner) with Gluon☆12Feb 22, 2019Updated 7 years ago
- Python Implementation of Reinforcement Learning: An Introduction☆14,742Aug 9, 2024Updated 2 years ago
- discrete gate sizing☆14Nov 23, 2020Updated 5 years ago
- Keras solution to the bAbI tasks using recurrent neural networks - merged as an example into Keras mainline☆33Aug 5, 2015Updated 11 years ago
- (TPAMI) Human-guided Reinforcement Learning with Sim-to-real Transfer for Autonomous Navigation☆26Sep 18, 2023Updated 2 years ago
- Ranking Policy Gradient☆23Nov 27, 2019Updated 6 years ago
- 中文整理的强化学习资料(Reinforcement Learning)☆2,189Apr 30, 2020Updated 6 years ago
- [KDD 2026] Official implementation of "FaST: Efficient and Effective Long-Horizon Forecasting for Large-Scale Spatial-Temporal Graphs via…☆17Jun 1, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implementation of CoDAIL in the ICLR 2020 paper <Multi-Agent Interactions Modeling with Correlated Policies>☆19Jun 17, 2021Updated 5 years ago
- ☆37Oct 20, 2020Updated 5 years ago
- Sparse Convex Optimization Toolkit (SCOT)☆13Feb 5, 2024Updated 2 years ago
- Navigation agent with Bayesian relational memory in the House3D environment☆30Sep 13, 2019Updated 6 years ago
- PyTorch 中文文档☆14Apr 6, 2018Updated 8 years ago
- A Spiking Multi-Layer Perceptron☆33Sep 5, 2017Updated 8 years ago
- machine learning trading system using random decision tree to train the technical indicators☆10Apr 11, 2017Updated 9 years ago
- 这是一个学习强化学习基础原理的仓库,主要包括了《深入浅出强化学习原理入门》书中一些例子和课后作业的代码☆272Dec 4, 2018Updated 7 years ago
- Multimodal Deep Q-Network (MDQN) for modelling human-like social intelligence.☆14Feb 23, 2017Updated 9 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Reinforcement Learning, Tutorials in Chinese☆11Jun 9, 2018Updated 8 years ago
- This project is focus on stock prediction,our goal is implementing one trading framework using DRL with LSTM.☆11Jun 1, 2018Updated 8 years ago
- This repo is using imitating learning to optimize portfolio. The code was derived from https://github.com/vermouth1992/drl-portfolio-mana…☆11May 16, 2019Updated 7 years ago
- ZERO is a modular C++ library interfacing Mathematical Programming and Game Theory.☆14Mar 12, 2024Updated 2 years ago
- Machine Learning written in TypeScript (to replace learn4js)☆10Apr 11, 2018Updated 8 years ago
- Data Types a la carte from PureScript -> JavaScript☆13Apr 19, 2017Updated 9 years ago
- Matlab Implementation of Autonomous Cross-Domain Knowledge Transfer in Lifelong Policy Gradient Reinforcement Learning☆17Mar 14, 2017Updated 9 years ago
- Benchmarking script for MindtPy solvers in the pyomo framework solving MINLP instance. This was created as part of my bachelor thesis dur…☆12Aug 13, 2020Updated 5 years ago
- ☆119Jul 9, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is a simple experiment designed to uncover which technical indicators are the most important.☆13Jan 11, 2019Updated 7 years ago
- Visual Navigation with Spatial Attention☆37Jan 2, 2025Updated last year
- Original implementation of QA4IE☆25Jul 28, 2021Updated 5 years ago
- Implement BinaryNet of CNN with chainer☆11May 5, 2016Updated 10 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- An open library of Generalized Disjunctive Programming (GDP) models☆13May 14, 2026Updated 2 months ago
- A implement of PGGAN for tensorflow version(progressive growing GANs for improved quality, stability and variation)☆14Dec 7, 2017Updated 8 years ago