《强化学习中 的数学原理》笔记-个人学习的思考和补充
☆114Jun 11, 2026Updated 2 months ago
Alternatives and similar repositories for Mathematical-Foundations-of-Reinforcement-Learning-Notes
Users that are interested in Mathematical-Foundations-of-Reinforcement-Learning-Notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 强化学习原理 + 强化学习代码实现 + 强化学习框架 + 强化学习论文☆35Aug 22, 2026Updated last week
- ☆33Jun 7, 2025Updated last year
- Multidigraph learning (MDGL) for training recurrent spiking neural networks☆14Dec 18, 2021Updated 4 years ago
- ☆21Feb 28, 2026Updated 6 months ago
- [ICML'26] Every Step Counts: Decoding Trajectories as Authorship Fingerprints of dLLMs☆35Oct 8, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Jan 6, 2025Updated last year
- ☆24Apr 13, 2026Updated 4 months ago
- Projects from basic algorithms to MARL. Implements MADDPG,MATD3,MA/HAPPO in Predator-Prey pursuit games with PettingZoo MPE environments.☆437Aug 10, 2026Updated 3 weeks ago
- Official repo for NeurIPS 2025 poster: Unveiling the Spatial-temporal Effective Receptive Fields of Spiking Neural Networks☆17Jul 30, 2026Updated last month
- VLM benchmarks for robot manipulation tasks☆23Apr 30, 2025Updated last year
- OpenVLA for AIRBOT☆17Aug 15, 2024Updated 2 years ago
- ☆14May 13, 2022Updated 4 years ago
- Metapackage that contains odometry nodes☆13Jul 1, 2026Updated last month
- 极不平衡样本下的预测☆40Oct 28, 2025Updated 10 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆14Jan 14, 2026Updated 7 months ago
- ☆24Mar 8, 2024Updated 2 years ago
- ☆303Jan 2, 2026Updated 7 months ago
- ☆18Oct 27, 2024Updated last year
- Exam & Example☆33May 1, 2025Updated last year
- The reinforcement learning training code for AgiBot X1.☆15Jan 15, 2025Updated last year
- A helper node to run KISS-ICP on ROS for those who are still obsessed with the dataset☆14May 25, 2023Updated 3 years ago
- This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."☆17,614Aug 10, 2026Updated 3 weeks ago
- Official code for "Efficient and robust temporal processing with neural oscillations modulated spiking neural networks" (Nature Communica…☆16Jul 8, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A nice framework with Fast-LIO2 as Frontend and a specific SLAM which applying scan_context++ as loop detection as Backend, collaborating…☆15Feb 13, 2025Updated last year
- 基于 Java 的动态验证框架,配置+脚本+热更新的方法来实现业务逻辑验证/业务规则获取☆19Jan 3, 2026Updated 7 months ago
- World Models That Know When They Don't Know: Controllable Video Generation with Calibrated Uncertainty☆25Dec 8, 2025Updated 8 months ago
- This repository is the implementation of "A Lightweight Spiking Neural Network for EEG-Based Motor Imagery Classification".☆16Jun 7, 2025Updated last year
- PyTorch Implementation of Online Training of Spiking Recurrent Neural Networks with Phase-Change Memory Synapses☆21Sep 25, 2021Updated 4 years ago
- ☆49Oct 27, 2025Updated 10 months ago
- [CVPR 2025] Lifelong Knowledge Editing for Vision Language Models with Low-Rank Mixture-of-Experts☆26Jun 22, 2025Updated last year
- ☆16May 29, 2025Updated last year
- SLAM homework based on LVI-SAM with BoW3D and Scan Context loop closure detection module adding.☆17Mar 20, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Cerebellar-SNN control of a Baxter robot. The repository includes EDLUT simulator source code, the ROS package to perform a closed-loop c…☆18Feb 4, 2022Updated 4 years ago
- ☆55Sep 26, 2025Updated 11 months ago
- A benchmark platform for robot grasping detection, integrating awesome projects and classic grasp algorithms.☆18Nov 18, 2023Updated 2 years ago
- ☆23Jun 14, 2025Updated last year
- Optimizing threshold and leak in LIF neurons with end-to-end backpropagation☆20Sep 10, 2021Updated 4 years ago
- loop closure test☆20Jan 20, 2025Updated last year
- A research-oriented collection of code, papers, and resources for legged robotics and model- and learning-based control.☆21Updated this week