David Silver【强化学习】Reinforcement Learning Course课件 该资源是David Silver的强化学习课程所对应的ppt课件。
☆16Apr 27, 2019Updated 7 years ago
Alternatives and similar repositories for DavidSilverRLPPT
Users that are interested in DavidSilverRLPPT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning (SASR)☆12Aug 26, 2025Updated last year
- 🍓 A toy object-oriented programming language written by rust☆17Apr 10, 2024Updated 2 years ago
- ☆14Oct 11, 2022Updated 3 years ago
- Code for paper: Reward Uncertainty for Exploration in Preference-based Reinforcement Learning☆15May 26, 2022Updated 4 years ago
- The official implementation of "Mind the Gap: Offline Policy Optimization for Imperfect Rewards" (ICLR2023)☆15Mar 3, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Spacecraft trajectory optimization using differential evolution☆14Sep 7, 2018Updated 8 years ago
- Official Codebase for TMLR 2023, Benchmarks and Algorithms for Offline Preference-Based Reward Learning☆20Dec 30, 2022Updated 3 years ago
- Listwise Reward Estimation for Offline Preference-based Reinforcement Learning (ICML 2024)☆18Jun 18, 2024Updated 2 years ago
- Implementation of Hash table for Nießner's Voxel Hashing method☆16Sep 2, 2015Updated 11 years ago
- Exports messages from topics in ROS bag files to CSV files. Matlab scripts then import the CSV files to Matlab workspaces.☆10Jun 15, 2016Updated 10 years ago
- a LiDAR-based Framework for Perception-aware Planning with Perturbation-induced Metric☆16Apr 18, 2025Updated last year
- Collision Avoidance using Buffered Voronoi Cell☆14Feb 10, 2017Updated 9 years ago
- This is the source codes of Recsys 2023 paper "Interpretable User Retention Modeling in Recommendation"☆26Mar 21, 2024Updated 2 years ago
- ☆11Nov 29, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 新版Mujoco学习记录☆21Apr 2, 2023Updated 3 years ago
- Program used to control and configure some of the ENSTA Bretagne UGVs, USVs, UUVs, UAVs used in WRSC, SAUC-E and euRathlon/ERL competitio…☆11Jun 16, 2026Updated 2 months ago
- Materials for the paper "Trajectory Replanning for Quadrotors Using Kinodynamic Search and Elastic Optimization"☆12Oct 16, 2017Updated 8 years ago
- Just example illustrates how the offline geographical maps capabilities can be added to Labview (using .Net control)☆13Apr 15, 2022Updated 4 years ago
- MATPOWER Extra - State estimation code contributed by Rui Bo.☆10May 14, 2024Updated 2 years ago
- Navigation control algorithms based on artificial vector fields☆12Aug 17, 2023Updated 3 years ago
- Some Multi-Agent Path Planning algorithms☆13Sep 27, 2020Updated 5 years ago
- A programmable servo capable of turning a defined 1 turn 360° (or more)☆13Dec 25, 2019Updated 6 years ago
- Reinforcement learning for load distribution in a decentralized Edge environment. This is the implementation of my Master's thesis projec…☆28Nov 28, 2024Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Code for the paper "Non-Linear Trajectory Optimization for Large Step-Ups: Application to the Humanoid Robot Atlas"☆19Mar 3, 2021Updated 5 years ago
- Pluggin and utils for viewing voxelgrids in RViz☆16May 10, 2021Updated 5 years ago
- Models rigged with muscles and environments which incorporate PyMuscle fatigable muscle models☆20Apr 6, 2019Updated 7 years ago
- Electric vehicle state of charge prediction using recurrent neural networks, submitted to UToronto's 2020 ProjectX competition representi…☆12Dec 3, 2020Updated 5 years ago
- 2018达观杯文本智能处理挑战赛☆16Nov 14, 2018Updated 7 years ago
- Multi-Agent Deep Reinforcement Learning for Collaborative Computation Offloading in Mobile Edge-Computing☆22May 29, 2025Updated last year
- Repository for the code examples of the 491/691 level "Autonomous Mobile Robot Design course"☆17Jul 21, 2016Updated 10 years ago
- Source code for my blog post tutorial about how to use deep learning on MR images.☆20Jan 12, 2024Updated 2 years ago
- Component-based structure for 6DOF AUV control☆14Oct 25, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 郑州大学(ZZU)简历 LaTeX 模板☆15Jan 11, 2026Updated 7 months ago
- ☆13Sep 1, 2019Updated 7 years ago
- BiC-MPPI: Goal-Pursuing, Sampling-Based Bidirectional Rollout Clustering Path Integral for Trajectory Optimization☆20Sep 25, 2024Updated last year
- ☆13Feb 24, 2023Updated 3 years ago
- Spatial-temporal Trajectory Planning for UAV Teach-and-Repeat☆16Jun 30, 2019Updated 7 years ago
- Automation of a decentralized swarm of autonomous mobile robots to perform various tasks through swarm-intelligence algorithms.☆17May 9, 2020Updated 6 years ago
- C++ Implementation of MPPI-IPDDP (Model Predictive Path Integral - Interior Point Differential Dynamic Programming) and Testing with MPPI…☆21Aug 13, 2024Updated 2 years ago