This repository contains my notes about deep reinforcement learning course in NJU.
☆14Jun 28, 2023Updated 3 years ago
Alternatives and similar repositories for deep_reinforcement_learning_notes
Users that are interested in deep_reinforcement_learning_notes are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Constructing a null model and measuring the nestedness of bipartite networks☆11Jul 23, 2020Updated 6 years ago
- ☆11Oct 7, 2022Updated 3 years ago
- ACeD: Scalable Data Availability Oracle☆14Oct 29, 2020Updated 5 years ago
- Implementation of BIMRL: Brain Inspired Meta Reinforcement Learning - Roozbeh Razavi et al. (IROS 2022)☆10Dec 1, 2022Updated 3 years ago
- ☆12Aug 30, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- TLA+ specification of Fast Flexible Paxos☆17Oct 9, 2020Updated 5 years ago
- TU Delft Blockchain Engineering course project on scale-out distributed ledger☆13Mar 4, 2018Updated 8 years ago
- Let TTY of the Linux kernel support UTF-8 (like CJKTTY☆11Aug 1, 2022Updated 3 years ago
- ☆12Apr 20, 2021Updated 5 years ago
- ☆20Oct 19, 2022Updated 3 years ago
- Computes affinity between two entities based on their co-occurrence☆30Jan 28, 2026Updated 5 months ago
- High-performance, high-scalability distributed computing with Erlang and Elixir.☆20Feb 8, 2023Updated 3 years ago
- TLA+ model checking guided testing for distributed systems☆19Feb 12, 2024Updated 2 years ago
- 树莓派+arduino+机械臂驱动器 通过摄像头控制机械臂下井字棋☆10Aug 28, 2019Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆16Feb 26, 2024Updated 2 years ago
- 计算传播学实验中心(中文网站)☆27Mar 2, 2026Updated 4 months ago
- scikit-learn compatible Python bindings for grf (generalized random forests) C++ random forest library☆34May 7, 2022Updated 4 years ago
- ☆14May 31, 2022Updated 4 years ago
- Adversarial Soft Advantage Fitting: Imitation Learning without Policy Optimization☆16Dec 10, 2020Updated 5 years ago
- Cryptid sdfix☆15Jul 3, 2025Updated last year
- ☆16Apr 12, 2021Updated 5 years ago
- Packet-level simulation code to model Opera and other networks from the 2020 NSDI paper "Expanding across time to deliver bandwidth effic…☆15Jun 10, 2020Updated 6 years ago
- Code to implement Maximum Entropy Deep Inverse Reinforcement Learning.☆14Jul 3, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 2025] "Understanding Constraint Inference in Safety-Critical Inverse Reinforcement Learning"☆16Nov 30, 2025Updated 7 months ago
- ☆15Sep 9, 2020Updated 5 years ago
- ☆37Oct 21, 2020Updated 5 years ago
- ☆14Sep 5, 2025Updated 10 months ago
- A simulation infrastructure for data center systems☆26Jul 17, 2012Updated 14 years ago
- ICML'2024: Q-value Regularized Transformer for Offline Reinforcement Learning☆38Dec 30, 2024Updated last year
- ☆35Mar 20, 2023Updated 3 years ago
- Code for "DrS: Learning Reusable Dense Rewards for Multi-Stage Tasks"☆22Apr 26, 2024Updated 2 years ago
- Code for paper "Learning Multimodal Transition Dynamics for Model-Based Reinforcement Learning".☆35May 24, 2018Updated 8 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Superfast CUDA implementation of Word2Vec and Latent Dirichlet Allocation (LDA)☆45Feb 25, 2021Updated 5 years ago
- The source for "Compiling with Dependent Types" (my dissertation)☆30May 10, 2022Updated 4 years ago
- Formal verification of the Algorand consensus protocol☆27Nov 20, 2022Updated 3 years ago
- Rollback protection for confidential services☆35Dec 21, 2025Updated 7 months ago
- 研究生论文☆15Oct 23, 2018Updated 7 years ago
- 极简爬虫工作流☆43May 22, 2023Updated 3 years ago
- TLA+ specifications accompanying paper: Automated Validation of State-Based Client-Centric Isolation with TLA+. (https://doi.org/10.1007/…☆27Feb 26, 2024Updated 2 years ago