这是我的强化学习笔记
☆228Jan 27, 2026Updated 5 months ago
Alternatives and similar repositories for Reinforcement-Learning-Study-Note
Users that are interested in Reinforcement-Learning-Study-Note are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 这是我的深度强化学习的学习笔记与总结☆85Mar 18, 2026Updated 4 months ago
- This is the homepage of a new book entitled "Mathematical Foundations of Reinforcement Learning."☆17,184May 26, 2026Updated last month
- Source code for ComNet paper: Satellite multi-beam multicast support for an efficient community-based CDN☆10Jul 26, 2022Updated 3 years ago
- ☆24Feb 27, 2026Updated 4 months ago
- to remove dynamic object in 3D pointcloud map☆17May 13, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆21Dec 10, 2024Updated last year
- 中国人民大学 YOJ 题库☆12Jun 9, 2022Updated 4 years ago
- standalone node and matlab wrapper for teb-planner package☆13Jul 1, 2021Updated 5 years ago
- ☆12Aug 30, 2024Updated last year
- Future Technologies Conference 2025 - MULTIMODAL EMOTION RECOGNITION AND SENTIMENT ANALYSIS IN MULTI-PARTY CONVERSATION CONTEXTS☆14Sep 12, 2024Updated last year
- Projects from basic algorithms to MARL. Implements MADDPG,MATD3,MA/HAPPO in Predator-Prey pursuit games with PettingZoo MPE environments.☆428Apr 8, 2026Updated 3 months ago
- EM algorithm: Gibbs sampling incorporated with Metropolis-Hastings step and computed the posterior mode through Louis' method☆16Nov 28, 2018Updated 7 years ago
- 现代化python入门教程☆43Nov 3, 2025Updated 8 months ago
- Xbotics 社区具身智能学习指南:我们把“具身综述→学习路线→仿真学习→开源实物→人物访谈→公司图谱”串起来,帮助新手和实战者快速定位路径、落地项目与参与开源。☆1,108Jun 28, 2026Updated 3 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆11May 24, 2023Updated 3 years ago
- The code about “LABNet: A Lightweight Attentive Beamforming Network for Ad-hoc Multichannel Microphone Invariant Real-Time Speech Enhance…☆49Oct 10, 2025Updated 9 months ago
- Implementation of RRT-Star to manipulate objects in ROS Gazebo with a 6-DOF robotic arm☆17Feb 11, 2024Updated 2 years ago
- Camera-LIDAR Fusion Framework for detection and tracking.☆16Sep 7, 2024Updated last year
- 基于 MAPPO (Multi-Agent Proximal Policy Optimization) 深度强化学习的多无人机三维协同编队、动态避障算法☆15Mar 27, 2026Updated 3 months ago
- 基于 LeRobot 和 MuJoCo 的机器人学习教程,包含 ACT、pi0、SmolVLA 模型的完整复现:数据采集、训练与部署。☆15Apr 26, 2026Updated 2 months ago
- 记录斯坦福公开课EE263的学习资料以及笔记。☆17Aug 29, 2019Updated 6 years ago
- Dynamic System Identification Toolbox☆13Jul 13, 2017Updated 9 years ago
- [IEEE TSIPN' 2022] "Scalable Perception-Action-Communication Loops with Convolutional and Graph Neural Networks", by Ting-Kuei Hu, Fernan…☆16Feb 4, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- 科研日志模板☆16May 8, 2025Updated last year
- Robot Agent using VLLMs to make long horizon plans. Isaac Sim, BEHAVIOR, Robosuite☆12Aug 31, 2025Updated 10 months ago
- Behavior Injection: Preparing Language Models for Reinforcement Learning (NeurIPS 2025)☆17Jul 1, 2025Updated last year
- clear single-file JAX implementations of common RL algorithms☆15Sep 5, 2021Updated 4 years ago
- A collection of tools I created related to the molecular simulations package RASPA.☆12Dec 4, 2023Updated 2 years ago
- Official repo for ICML 2025 paper "RollingQ: Reviving the Cooperation Dynamics in Multimodal Transformer"☆17Jun 21, 2025Updated last year
- If you can read ~100 lines of Python, you understand Skills.☆19Mar 16, 2026Updated 4 months ago
- Xbotics 灵巧手平台:多指操作基础与进阶☆27Jun 9, 2026Updated last month
- MATLAB code for component-informed data-center power-delivery modeling and power-system resonance analysis.☆18Jun 26, 2026Updated 3 weeks ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICLR 2026 Oral] MomaGraph: State-Aware Unified Scene Graphs with Vision-Language Models for Embodied Task Planning☆56Mar 4, 2026Updated 4 months ago
- ☆15Jul 24, 2024Updated last year
- A 2D simulation in Pygame of the paper "Randomized Kinodynamic Planning" by Steven M. LaValle, and James J. Kuffner, Jr.☆27Apr 18, 2023Updated 3 years ago
- Unofficial repository for the Power Systems Analysis Toolbox by Federico Milano☆17Dec 8, 2023Updated 2 years ago
- Code for "Strengthened and Faster Linear Approximation to Joint Chance Constraints with Wasserstein Ambiguity", INFORMS Journal of Comput…☆19Feb 26, 2025Updated last year
- Modification of RASPA2 code for GC-TMMC simulations☆13Apr 18, 2024Updated 2 years ago
- OpenAI Gym environments for spacecraft operations problems.☆14Mar 25, 2021Updated 5 years ago