Exercise Solutions for Reinforcement Learning: An Introduction [2nd Edition]
☆16Jul 17, 2020Updated 6 years ago
Alternatives and similar repositories for rlai-exercises
Users that are interested in rlai-exercises are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Plan time-optimal paths with both speed and turn-rate controls☆10May 15, 2021Updated 5 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- RefTeacher is a strong baseline method for Semi-Supervised Referring Expression Comprehension.☆14May 26, 2023Updated 3 years ago
- Code for submission to 2024 submission to Automatica titled "Closed-loop Data-enabled Predictive Control and its equivalence with Closed-…☆14Sep 26, 2024Updated 2 years ago
- ☆21Aug 24, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Benchmark result of different RL algorithms on MetaDrive environments, including Multi-agent RL (IPPO, centralized critics, CoPO).☆16Oct 25, 2022Updated 3 years ago
- ☆18Jun 18, 2024Updated 2 years ago
- 网易云音乐命令行版本,排行榜,搜索,精选歌单,登录,DJ节目,快速打碟,本地收藏歌单☆19May 19, 2016Updated 10 years ago
- ☆18Oct 6, 2021Updated 4 years ago
- Code/data of the paper "Hand-Object Contact Prediction via Motion-Based Pseudo-Labeling and Guided Progressive Label Correction" (BMVC202…☆18Oct 22, 2021Updated 4 years ago
- [ICCV 2025] A Benchmark for Multi-Step Reasoning in Long Narrative Videos☆28Jun 4, 2026Updated 3 months ago
- Support library for the MaskRCNN masks extracted on EPIC-KITCHENS-100☆14Dec 1, 2020Updated 5 years ago
- PyTorch implementation for "Generative Modeling on Manifolds Through Mixture of Riemannian Diffusion Processes" (ICML 2024).☆13Jul 21, 2024Updated 2 years ago
- 基于socket tcp通信,使用tkinter做客户端界面;一个多人同时在线的聊天系统;python课程设计☆24Dec 13, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- fixed wing uav model test☆12Jan 15, 2016Updated 10 years ago
- Kuka Reacher Reinforcement Learning Sim2Real Environment for Omniverse Isaac Gym/Sim☆21Nov 22, 2023Updated 2 years ago
- Solution for Taxi env using HRL (Hierarchical reinforcement learning) (2018)☆21Nov 3, 2019Updated 6 years ago
- code for the ddp tutorial☆33Apr 9, 2022Updated 4 years ago
- Airplanes War game, based on Unity 3D game engine.☆15Oct 30, 2020Updated 5 years ago
- ☆18Mar 19, 2019Updated 7 years ago
- My Solutions to Sutton and Barto exercises, 2nd edition☆14Apr 27, 2018Updated 8 years ago
- ☆14Mar 11, 2026Updated 6 months ago
- Multi-Agent Reinforcement Learning☆11Jun 16, 2020Updated 6 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR 2023] Official code release of Cafi-Net: Self-Supervised Learning of Pose-Canonicalized Neural Fields☆15Jul 14, 2023Updated 3 years ago
- Implementation of Proximal Policy Optimization algorithm on a custom Unity environment.☆17Feb 3, 2022Updated 4 years ago
- Score-Based Diffusion Policy Compatible with Reinforcement Learning via Optimal Transport☆15Feb 26, 2025Updated last year
- ☆27Dec 7, 2019Updated 6 years ago
- ☆12Sep 7, 2024Updated 2 years ago
- Python D* Lite☆29Sep 18, 2018Updated 8 years ago
- Code for Learned Thresholds Token Merging and Pruning for Vision Transformers (LTMP). A technique to reduce the size of Vision Transforme…☆17Nov 24, 2024Updated last year
- academic pages for c2 group☆11Apr 11, 2023Updated 3 years ago
- ☆19Dec 23, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- picshare☆16Jan 15, 2022Updated 4 years ago
- ☆24Aug 9, 2022Updated 4 years ago
- Code for "PUMA: Deep Metric Imitation Learning for Stable Motion Primitives"☆18Oct 1, 2024Updated last year
- Simulate Grasp Dataset Generation (未完成版)主要实现的功能是,使用Antipodal算法对虚拟的物理环境中的Mesh模型进行6-Dof抓取采样☆14Jun 14, 2022Updated 4 years ago
- Results reproductions & comparisons between OpenSpiel implementations, associated paper & originating works☆18Mar 2, 2021Updated 5 years ago
- PyBullet simulator for Franka Emika Panda☆15Jul 30, 2020Updated 6 years ago
- Code accompanying the paper "TamedPUMA: safe and stable imitation learning with geometric fabrics" (L4DC 2025)☆19Aug 18, 2026Updated last month