CS285课程笔记
☆25Jan 19, 2020Updated 6 years ago
Alternatives and similar repositories for CS285_Note_CN
Users that are interested in CS285_Note_CN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Papers related to Legged locomotion☆17Oct 16, 2024Updated last year
- My solution to assignments for Berkeley CS 285: Deep Reinforcement Learning, Decision Making, and Control.☆16Mar 19, 2025Updated last year
- ☆10Dec 18, 2023Updated 2 years ago
- Implementation of "Active Exploration for Inverse Reinforcement Learning (AceIRL), NeurIPS 2022.☆14Oct 12, 2022Updated 3 years ago
- ☆17Aug 12, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Soft Actor-Critic implementation with SOTA model-free extension (REDQ) and SOTA model-based extension (MBPO).☆15Feb 21, 2021Updated 5 years ago
- [NeurIPS 2023] MoVie: Visual Model-Based Policy Adaptation for View Generalization☆12Sep 22, 2023Updated 2 years ago
- PyTorch implementation of the paper Overcoming Exploration in Reinforcement Learning with Demonstrations in surgical robot manipulation t…☆12Aug 21, 2022Updated 3 years ago
- cmdr cxx version, a C++17/20 header-only command-line parser with hierarchical config data manager here☆18Jul 14, 2026Updated last month
- 基于蒙特卡洛树搜索算法编写的黑白棋AI算法☆12Mar 28, 2022Updated 4 years ago
- Othello AI | Monte Carlo tree search☆16Sep 9, 2018Updated 7 years ago
- BeCL: Behavior Contrastive Learning for Unsupervised Skill Discovery.☆23May 11, 2023Updated 3 years ago
- [NeurIPS 2021] World modelling and action learning using a contrastive formulation of the active inference framework, for reaching visual…☆15Jan 22, 2024Updated 2 years ago
- Mod Source for BOTATO, brotato auto-battler mod☆22Aug 6, 2026Updated last week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 2D Iterative Learning Control with Deep Reinforcement Learning Compensation for the Non-repetitive Batch Processes☆11Mar 4, 2025Updated last year
- 哈工大2022春 高级算法(本科)课程实验☆14May 27, 2022Updated 4 years ago
- [ EMNLP 2025 Main ] Enhancing Efficiency and Exploration in Reinforcement Learning for LLMs☆18Nov 7, 2025Updated 9 months ago
- This is a repository of DQN and its variants implementation in PyTorch based on the original papar.☆13Nov 18, 2019Updated 6 years ago
- ☆17Nov 18, 2024Updated last year
- ☆44Updated this week
- Safe Model-based Reinforcement Learning with Robust Cross-Entropy Method☆65Mar 24, 2023Updated 3 years ago
- Code for IEEE transactions on neural networks and learning system☆13Jun 18, 2021Updated 5 years ago
- ☆24Sep 30, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SRSA: Skill Retrieval and Adaptation for Robotic Assembly Tasks☆18Mar 25, 2026Updated 4 months ago
- A C++ library to benchmark inverted indexes.☆23Aug 4, 2020Updated 6 years ago
- ☆26Jan 26, 2024Updated 2 years ago
- RePO: Replay-Enhanced Policy Optimization☆24Jun 12, 2025Updated last year
- Build DDPG models and test on stock market☆22Nov 19, 2018Updated 7 years ago
- Image Deduplication in Python☆23May 16, 2020Updated 6 years ago
- An implementation of TRPO with GAE in PyTorch☆16Jul 22, 2023Updated 3 years ago
- Official code for ICLR 2024 paper, SEABO: A Simple Search-Based Method for Offline Imitation Learning☆12Jan 19, 2024Updated 2 years ago
- Greenhouse Energy Simulation☆24Oct 5, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- fork Yunjey Choi/ [@yunjey](https://github.com/yunjey)☆23Jan 23, 2019Updated 7 years ago
- USAD model on UCR Time Series Anomaly Archive☆15Oct 22, 2021Updated 4 years ago
- [NeurIPS 2023] Conformal Prediction for Uncertainty-Aware Planning with Diffusion Dynamics Model☆20Dec 9, 2023Updated 2 years ago
- Massively multiagent reinforcement learning in a slither.io like environment☆24Dec 8, 2022Updated 3 years ago
- Project for Course : Reinforcement Learning☆16Apr 29, 2020Updated 6 years ago
- Code for MOBILE: Model-Bellman Inconsistency Penalized Offline Policy Optimization☆22Apr 17, 2024Updated 2 years ago
- Official code for Cross-Domain Policy Adaptation by Capturing Representation Mismatch (ICML 2024)☆15Aug 15, 2025Updated last year