Exercise Solutions for Reinforcement Learning: An Introduction [2nd Edition]
☆16Jul 17, 2020Updated 6 years ago
Alternatives and similar repositories for rlai-exercises
Users that are interested in rlai-exercises are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Plan time-optimal paths with both speed and turn-rate controls☆10May 15, 2021Updated 5 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- [ICLR 2026] Block-wise Adaptive Caching for Accelerating Diffusion Policy☆18Jan 27, 2026Updated 6 months ago
- Code for submission to 2024 submission to Automatica titled "Closed-loop Data-enabled Predictive Control and its equivalence with Closed-…☆14Sep 26, 2024Updated last year
- ☆19Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Benchmark result of different RL algorithms on MetaDrive environments, including Multi-agent RL (IPPO, centralized critics, CoPO).☆16Oct 25, 2022Updated 3 years ago
- This is a collection of Matlab functions that are useful in the development of target tracking algorithms.☆15Sep 3, 2015Updated 10 years ago
- GCN CAV☆14Mar 29, 2021Updated 5 years ago
- Code/data of the paper "Hand-Object Contact Prediction via Motion-Based Pseudo-Labeling and Guided Progressive Label Correction" (BMVC202…☆17Oct 22, 2021Updated 4 years ago
- fixed wing uav model test☆12Jan 15, 2016Updated 10 years ago
- Solution for Taxi env using HRL (Hierarchical reinforcement learning) (2018)☆21Nov 3, 2019Updated 6 years ago
- Airplanes War game, based on Unity 3D game engine.☆15Oct 30, 2020Updated 5 years ago
- ☆18Mar 19, 2019Updated 7 years ago
- 哈工大深圳 校园网全自动登录☆10Feb 7, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- PPDDL plan evalutation simulator☆15Dec 30, 2019Updated 6 years ago
- ☆14Mar 11, 2026Updated 5 months ago
- Multi-Agent Reinforcement Learning☆11Jun 16, 2020Updated 6 years ago
- [CVPR 2023] Official code release of Cafi-Net: Self-Supervised Learning of Pose-Canonicalized Neural Fields☆15Jul 14, 2023Updated 3 years ago
- Implementation of Proximal Policy Optimization algorithm on a custom Unity environment.☆17Feb 3, 2022Updated 4 years ago
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- This is an official implementation for "ManipForce: Force-Guided Policy Learning with Frequency-Aware Representation for Contact-Rich Man…☆14Apr 23, 2026Updated 3 months ago
- Code for Learned Thresholds Token Merging and Pruning for Vision Transformers (LTMP). A technique to reduce the size of Vision Transforme…☆17Nov 24, 2024Updated last year
- Non-official implementation of paper "In-context Reinforcement Learning with Algorithm Distillation"☆13Aug 15, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆19Dec 23, 2024Updated last year
- picshare☆16Jan 15, 2022Updated 4 years ago
- ☆24Aug 9, 2022Updated 4 years ago
- Official implementation of the NeurIPS 25 paper of Riemannian Consistency Model (RCM) for few-step generation on Riemannian manifolds.☆17Nov 2, 2025Updated 9 months ago
- Hydronautics team ROS based framework for autonomous underwater vehicles (AUV)☆25Apr 5, 2026Updated 4 months ago
- Results reproductions & comparisons between OpenSpiel implementations, associated paper & originating works☆18Mar 2, 2021Updated 5 years ago
- pybullet grasping with time contrastive network embeddings☆22Jun 18, 2019Updated 7 years ago
- PyBullet simulator for Franka Emika Panda☆15Jul 30, 2020Updated 6 years ago
- object detection on DIOR☆22Jan 29, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code accompanying the paper "TamedPUMA: safe and stable imitation learning with geometric fabrics" (L4DC 2025)☆19Updated this week
- DARP+STC algorithm for mCPP problem☆17Mar 29, 2019Updated 7 years ago
- DBN++ Data Structures and Algorithms in C++ for Dynamic Bayesian Networks☆18Feb 5, 2016Updated 10 years ago
- A short conceptual replication of "Prefrontal cortex as a meta-reinforcement learning system" in Jax.☆19Feb 27, 2023Updated 3 years ago
- Codex plugin for local Obsidian workflows through the official desktop CLI.☆23Apr 22, 2026Updated 3 months ago
- 数学建模:算法与编程实现书籍配套资料☆19Jan 10, 2023Updated 3 years ago
- ☆24Jul 29, 2026Updated 2 weeks ago