This is the repository hosting the R scripts of the book "Mathematical Foundations of Reinforcement Learning" written by Yujun at Jiangxi Normal University.
☆26Dec 3, 2023Updated 2 years ago
Alternatives and similar repositories for Code-Mathmatical-Foundation-of-Reinforcement-Learning
Users that are interested in Code-Mathmatical-Foundation-of-Reinforcement-Learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DQN algorithm by Pytorch - a simple maze game☆38Apr 14, 2023Updated 3 years ago
- CS 188 Project 3☆11Mar 5, 2018Updated 8 years ago
- PyTorch implementation of PtrNet to solve sorting problem.☆12Dec 19, 2017Updated 8 years ago
- 该工具包主要完成在项目内根据提供的实体类包,自动生成spring mybatis,所需要的service层接口与实现,数据库表的创建包括主键,描述,长度等的设置,数据库操作接口与对应xml文件,支持jar与maven☆10Dec 23, 2022Updated 3 years ago
- ☆12Apr 21, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆13Jun 26, 2020Updated 6 years ago
- NeurIPS 2025: Improving Monte Carlo Tree Search for Symbolic Regression☆17Jun 7, 2026Updated 2 months ago
- Encode-attend-navigate unofficial Pytorch implementation☆12Oct 1, 2024Updated last year
- MATLAB implementation of DQN for a navigation environment☆13Aug 13, 2020Updated 6 years ago
- This project applies Monte Carlo Tree Search (MCTS) to a simple grid world.☆10May 30, 2018Updated 8 years ago
- Solving VRPC with column generation and branch and price for fun and profit☆13Mar 27, 2023Updated 3 years ago
- This repository contains the implementation of a Deep Deterministic Policy Gradient (DDPG) algorithm applied to solve the Reacher environ…☆12Apr 8, 2023Updated 3 years ago
- Reinforcement Learning Recommender System suggesting relevant scientific services to appropriate researchers☆11Aug 29, 2024Updated last year
- The implementation of STAR-HiT.☆11Oct 18, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Autoware V2X module with Zenoh☆14Jun 25, 2026Updated last month
- Vehicle Route Optimization with Reinforcement Learning (SARSA and Q_Learning) for Final Year Project☆14Aug 24, 2023Updated 2 years ago
- MagnetoPyElastica, an extension of PyElastica, is an open-source project for simulating magnetic Cosserat rods interacting with external …☆17Jul 27, 2024Updated 2 years ago
- This is the Github repository containing the code for the Context-Aware Sequential Recommendation project for the Information Retrieval 2…☆11Mar 24, 2023Updated 3 years ago
- ☆15Mar 24, 2024Updated 2 years ago
- ☆17Jul 15, 2026Updated last month
- Single robot path planning algorithms implemented in MATLAB. Including heuristic search and incremental heuristic search methods. A*, LPA…☆23Jul 7, 2023Updated 3 years ago
- ☆18Dec 5, 2017Updated 8 years ago
- ☆11May 5, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code repository for the book Feature engineering with Feature-engine☆15Apr 27, 2024Updated 2 years ago
- Machine Learning, Python☆10Dec 20, 2023Updated 2 years ago
- This is the official repository of the AI for TSP competition at IJCAI 2021☆28Nov 22, 2022Updated 3 years ago
- Lanelet2 map tile generator tool for Autoware dynamic lanelet2 map loading.☆12Jun 24, 2024Updated 2 years ago
- ☆28Jun 26, 2026Updated last month
- Created a computer vision pipeline to detect and classify cars as SUVs or sedans using transfer learning on Mobilenet and object detectio…☆14Mar 27, 2021Updated 5 years ago
- 深蓝学院 - 高飞 - 运动规划课程作业☆26Feb 16, 2022Updated 4 years ago
- Accompanying code for the text: Maniezzo, Vittorio, Boschetti, Marco Antonio, Stützle, Thomas "Matheuristics, algorithms and implementati…☆18Apr 5, 2024Updated 2 years ago
- 国科大疫情自动打卡脚本,根据上次提交记录生成本次提交记录,可选择账号密码登录或cookies登录。☆15Sep 6, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A 3D Navigation algorithm combining DWA and A*, Made in MATLAB☆32Jul 2, 2023Updated 3 years ago
- A reinforcement learning environment for aircraft control using the JSBSim flight dynamics model☆14Nov 7, 2020Updated 5 years ago
- ☆303Jan 2, 2026Updated 7 months ago
- Tiansuan Experiment Platform aims to enable the global academic community to conduct experiments on real satellites and to evaluate pract…☆14Jul 19, 2024Updated 2 years ago
- This repository accompanies our research paper titled "An LLM-based Recommender System Environment".☆17Jul 15, 2024Updated 2 years ago
- Implementation of Pareto Deep Q Networks in a multi-objective Gym Reinforcement Learning Environment☆18Jun 19, 2023Updated 3 years ago
- ☆18Mar 21, 2019Updated 7 years ago