Exercise Solutions for Reinforcement Learning: An Introduction [2nd Edition]
☆16Jul 17, 2020Updated 6 years ago
Alternatives and similar repositories for rlai-exercises
Users that are interested in rlai-exercises are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Plan time-optimal paths with both speed and turn-rate controls☆10May 15, 2021Updated 5 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- [ICLR 2026] Block-wise Adaptive Caching for Accelerating Diffusion Policy☆18Jan 27, 2026Updated 6 months ago
- Code for submission to 2024 submission to Automatica titled "Closed-loop Data-enabled Predictive Control and its equivalence with Closed-…☆14Sep 26, 2024Updated last year
- ☆19Jun 29, 2026Updated last month
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Benchmark result of different RL algorithms on MetaDrive environments, including Multi-agent RL (IPPO, centralized critics, CoPO).☆16Oct 25, 2022Updated 3 years ago
- This is a collection of Matlab functions that are useful in the development of target tracking algorithms.☆15Sep 3, 2015Updated 10 years ago
- ☆18Oct 6, 2021Updated 4 years ago
- GCN CAV☆14Mar 29, 2021Updated 5 years ago
- Code/data of the paper "Hand-Object Contact Prediction via Motion-Based Pseudo-Labeling and Guided Progressive Label Correction" (BMVC202…☆17Oct 22, 2021Updated 4 years ago
- Support library for the MaskRCNN masks extracted on EPIC-KITCHENS-100☆14Dec 1, 2020Updated 5 years ago
- ☆11May 15, 2024Updated 2 years ago
- fixed wing uav model test☆12Jan 15, 2016Updated 10 years ago
- Kuka Reacher Reinforcement Learning Sim2Real Environment for Omniverse Isaac Gym/Sim☆21Nov 22, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The official repo for [ACM CSUR'24] "Empowering Agrifood System with Artificial Intelligence: A Survey of the Progress, Challenges and Op…☆12Dec 6, 2024Updated last year
- Solution for Taxi env using HRL (Hierarchical reinforcement learning) (2018)☆21Nov 3, 2019Updated 6 years ago
- Airplanes War game, based on Unity 3D game engine.☆15Oct 30, 2020Updated 5 years ago
- ☆18Mar 19, 2019Updated 7 years ago
- My Solutions to Sutton and Barto exercises, 2nd edition☆14Apr 27, 2018Updated 8 years ago
- PPDDL plan evalutation simulator☆15Dec 30, 2019Updated 6 years ago
- ☆14Mar 11, 2026Updated 4 months ago
- [CVPR 2023] Official code release of Cafi-Net: Self-Supervised Learning of Pose-Canonicalized Neural Fields☆15Jul 14, 2023Updated 3 years ago
- Implementation of Proximal Policy Optimization algorithm on a custom Unity environment.☆17Feb 3, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- ☆12Sep 7, 2024Updated last year
- ☆30Aug 20, 2021Updated 4 years ago
- Python D* Lite☆29Sep 18, 2018Updated 7 years ago
- This is an official implementation for "ManipForce: Force-Guided Policy Learning with Frequency-Aware Representation for Contact-Rich Man…☆15Apr 23, 2026Updated 3 months ago
- ☆20Jun 30, 2021Updated 5 years ago
- Code for Learned Thresholds Token Merging and Pruning for Vision Transformers (LTMP). A technique to reduce the size of Vision Transforme…☆17Nov 24, 2024Updated last year
- Non-official implementation of paper "In-context Reinforcement Learning with Algorithm Distillation"☆13Aug 15, 2024Updated last year
- Branch of JavaFF planner for PDDL2.1☆16Oct 18, 2018Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- In this repository you can find the resuls of the simulated evaluation of an innovative, optimized for real-life use, STC-based, multi-ro…☆15Jan 20, 2022Updated 4 years ago
- ☆19Dec 23, 2024Updated last year
- [NeurIPS 2022 Spotlight] Hand-Object Interaction Image Generation☆33Nov 29, 2022Updated 3 years ago
- Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks☆16Oct 2, 2025Updated 9 months ago
- ☆24Aug 9, 2022Updated 3 years ago
- Official implementation of the NeurIPS 25 paper of Riemannian Consistency Model (RCM) for few-step generation on Riemannian manifolds.☆17Nov 2, 2025Updated 8 months ago
- Code for "PUMA: Deep Metric Imitation Learning for Stable Motion Primitives"☆18Oct 1, 2024Updated last year