This repository contains my notes about deep reinforcement learning course in NJU.
☆14Jun 28, 2023Updated 3 years ago
Alternatives and similar repositories for deep_reinforcement_learning_notes
Users that are interested in deep_reinforcement_learning_notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Notes for Distributionally Robust Optimization (DRO) 分布鲁棒优化学习笔记☆55Mar 29, 2023Updated 3 years ago
- ☆11Oct 7, 2022Updated 3 years ago
- This repository contains code snippets discussed in 15-440, lecture 5 (given on 1/28/2014).☆11Jan 30, 2014Updated 12 years ago
- Official code for ACT: Empowering Decision Transformer with Dynamic Programming via Advantage Conditioning (AAAI'24)☆16Feb 10, 2024Updated 2 years ago
- Implementation of BIMRL: Brain Inspired Meta Reinforcement Learning - Roozbeh Razavi et al. (IROS 2022)☆10Dec 1, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆12Aug 30, 2024Updated 2 years ago
- 域渗透脑图中文翻译版☆12Jan 26, 2022Updated 4 years ago
- This is a project based on machine learning and deep learning method for playing Gobang by controlling mechanical arm(利用机械臂下五子棋)☆13Apr 16, 2023Updated 3 years ago
- 经典推荐算法的python实现或使用,涵盖协同过滤、矩阵分解、gbdt+lr,以及Wide&Deep等深度推荐模型。☆10Feb 9, 2022Updated 4 years ago
- Packing Algorithm&LP Search&Learn to Pack(Undergraduate Research)☆35Aug 22, 2020Updated 6 years ago
- TLA+ specification of Fast Flexible Paxos☆17Oct 9, 2020Updated 5 years ago
- Code to reproduce paper results (or as close as possible, depending on data-availability). Each publication has a Jupyter notebook. Mostl…☆12Mar 8, 2024Updated 2 years ago
- TU Delft Blockchain Engineering course project on scale-out distributed ledger☆13Mar 4, 2018Updated 8 years ago
- BFT-Store is a Byzantine fault-tolerant storage engine☆14Sep 9, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Predicting the medal table of the Summer Games☆12Jul 6, 2023Updated 3 years ago
- ☆12Apr 20, 2021Updated 5 years ago
- 哔哩哔哩每日任务,每日任务升6级,直播心跳 小心心,天选时刻,赛事竞猜。支持Docker,青龙面板,以及各种云函数。☆16Jan 1, 2023Updated 3 years ago
- 树莓派+arduino+机械臂驱动器 通过摄像头控制机械臂下井字棋☆10Aug 28, 2019Updated 7 years ago
- High-performance, high-scalability distributed computing with Erlang and Elixir.☆20Feb 8, 2023Updated 3 years ago
- ☆16Feb 26, 2024Updated 2 years ago
- YOLO Series☆14Oct 20, 2023Updated 2 years ago
- A gesture recognition system powered by OpenPose, k-nearest neighbours, and local outlier factor.☆17Jun 7, 2021Updated 5 years ago
- ☆14May 31, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Cryptid sdfix☆16Jul 3, 2025Updated last year
- Adversarial Soft Advantage Fitting: Imitation Learning without Policy Optimization☆16Dec 10, 2020Updated 5 years ago
- ☆12May 22, 2018Updated 8 years ago
- ☆16Apr 12, 2021Updated 5 years ago
- PyTorch code for TAPAS-GMM.☆16Nov 21, 2024Updated last year
- Code to implement Maximum Entropy Deep Inverse Reinforcement Learning.☆14Jul 3, 2020Updated 6 years ago
- ZJU机器人学ROS移动机器人作业☆17Jun 27, 2020Updated 6 years ago
- [ICLR 2025] "Understanding Constraint Inference in Safety-Critical Inverse Reinforcement Learning"☆16Nov 30, 2025Updated 9 months ago
- 大工生存手册☆31Sep 14, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ICML'2024: Q-value Regularized Transformer for Offline Reinforcement Learning☆39Dec 30, 2024Updated last year
- Code for "DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks"☆22Apr 26, 2024Updated 2 years ago
- The source for "Compiling with Dependent Types" (my dissertation)☆30May 10, 2022Updated 4 years ago
- Rollback protection for confidential services☆35Aug 24, 2026Updated last month
- Lecture Code for Introduction to R Programming☆26Dec 1, 2020Updated 5 years ago
- A Ledger-backed Secure Key-Value store (LSKV), built on the Confidential Consortium Framework (CCF)☆38Feb 12, 2026Updated 7 months ago
- Reinforcment Learning accelerated Motion Planning Algorithm based on Frenetix and Frenetix_Motion_Planner☆80Feb 8, 2024Updated 2 years ago