斯坦福 cs234 强化学习中文讲义
☆215Jan 2, 2021Updated 5 years ago
Alternatives and similar repositories for stanford-cs234-notes-zh
Users that are interested in stanford-cs234-notes-zh are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 我的强化学习笔记和学习材料 still updating ... ...☆376Sep 27, 2025Updated last year
- 中文整理的强化学习资料(Reinforcement Learning)☆2,195Apr 30, 2020Updated 6 years ago
- Gradient descent algorithms for LQG control☆14Feb 20, 2022Updated 4 years ago
- 这是一个学习强化学习基础原理的仓库,主要包括了《深入浅出强化学习原理入门》书中一些例子和课后作业的代码☆272Dec 4, 2018Updated 7 years ago
- [NeurIPS'20] Code for the paper "Offline Imitation Learning with a Misspecified Simulator"☆12Nov 24, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [译] Gainlo 面试指南☆19Sep 17, 2020Updated 6 years ago
- Deep Reinforcement Learning Lab, a platform designed to make DRL technology and fun for everyone☆2,572Apr 11, 2022Updated 4 years ago
- Intro to Reinforcement Learning (强化学习纲要)☆3,602Jul 25, 2020Updated 6 years ago
- MATLAB code to produce results and figures in the paper "Stochastic Optimal Control of Pairs Trading Strategies with Absolute and Relativ…☆15Jun 1, 2018Updated 8 years ago
- 🕹️ CS234: Reinforcement Learning, Winter 2019 | YouTube videos 👉☆315Mar 25, 2023Updated 3 years ago
- ☆27Apr 22, 2024Updated 2 years ago
- Modelling bus-on-demand using SUMO and TraCI.☆19Apr 30, 2014Updated 12 years ago
- 《Reinforcement Learning: An Introduction》(第二版)中文翻译☆696Updated this week
- Some basic examples of playing with RL☆1,273Feb 18, 2026Updated 7 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 强化学习中文教程(蘑菇书🍄),在线阅读地址:https://datawhalechina.github.io/easy-rl/☆14,705Dec 30, 2025Updated 8 months ago
- ☆12May 14, 2024Updated 2 years ago
- Revisiting Peng's Q(lambda) for Modern Reinforcement Learning☆15Jul 23, 2021Updated 5 years ago
- 2 algorithms of optimal trade execution: 1) Dynamic Programming 2) Frank-Wolfe Algorithm (Python & C++)☆19Dec 11, 2019Updated 6 years ago
- [译] UCB DS100 数据科学的原理与技巧☆116Jan 2, 2021Updated 5 years ago
- Code for the paper {Pang, Bo, and Zhong-Ping Jiang. "Reinforcement Learning for Adaptive Optimal Stationary Control of Linear Stochastic …☆28Dec 5, 2021Updated 4 years ago
- Python Implementation of Reinforcement Learning: An Introduction☆14,765Aug 9, 2024Updated 2 years ago
- AI项目(强化学习、深度学习、计算机视觉、推荐系统、自然语言处理、机器导航、医学影像处理)☆95Aug 8, 2023Updated 3 years ago
- paper list in the area of reinforcenment learning for recommendation systems☆25Aug 4, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆25Aug 25, 2021Updated 5 years ago
- ☆12Mar 21, 2024Updated 2 years ago
- A translation of Reinforcement Learning: An Introduction☆114Aug 20, 2018Updated 8 years ago
- Statsmodels: Python中的统计建模与计量统计学类库,此为ApacheCN推出的中文版翻译。☆179Apr 21, 2021Updated 5 years ago
- Consistency Conditions for any two X-ray images.☆14Jun 14, 2026Updated 3 months ago
- [译] PythonBasics 中文系列教程☆25Jul 7, 2022Updated 4 years ago
- ☆38Mar 28, 2022Updated 4 years ago
- A Chinese learning note with python codes for Pattern Recognition and Machine Learning.☆33Aug 25, 2018Updated 8 years ago
- 本项目以一个可视化配置的、以AgentRL为核心的强化学习框架,实现30分钟上手AgentRL 编程。后续增加AgentRL和本地Agent、MCP、A2A相关特性。☆80Jul 9, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Un-official technical documents aimed at helping building Apollo auto-driving system on prototype cars☆10Jul 18, 2020Updated 6 years ago
- Simple tabular Q learning to solve the travelling salesman problem.☆10Jul 23, 2023Updated 3 years ago
- tensorflow实战练习,包括强化学习、推荐系统、nlp等☆7,028Sep 24, 2023Updated 3 years ago
- Ride Hailing Simulation - A data-driven approach to model a simulation environment☆13Feb 16, 2021Updated 5 years ago
- Android PopupWindows with an arrow indicator☆11Mar 18, 2016Updated 10 years ago
- ALSET Autonomous Vehicles: Model S: A toy RC robot with arm and tracked wheels. Model X: a toy RC Excavator. Both run the same software…☆12Aug 21, 2026Updated last month
- Implementations of LCA and RMQ data structures from "The LCA Problem Revisited"☆15Aug 19, 2014Updated 12 years ago