[动手学强化学习]系列,基于pytorch。
☆59Jun 2, 2021Updated 5 years ago
Alternatives and similar repositories for reinforcement_learning
Users that are interested in reinforcement_learning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Multi-objective application placement in fog computing using graph neural network-based reinforcement learning☆10Oct 20, 2025Updated 9 months ago
- ☆11Feb 28, 2022Updated 4 years ago
- dqn autoplay mario bros☆21Jul 24, 2017Updated 9 years ago
- pytorch implementation of DQN, NAF, DDPG☆13Jun 7, 2018Updated 8 years ago
- Repository for Robust Trajectory Optimization with Stochastic Complementarity☆12Dec 15, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Reinforcement Learning for Self Organization and Power Control of Two-Tier Heterogeneous Networks☆20Jan 18, 2020Updated 6 years ago
- Minimal UMI deployment environment for ARX5 robot arm☆23Feb 25, 2025Updated last year
- A.I. snake game build using pygame | pytorch | python☆17May 5, 2026Updated 2 months ago
- 预测-校正学习计算制导律☆13Jun 22, 2021Updated 5 years ago
- Python-based cross-platform tool for mining text data (html, transcript, problems) of edX MOOCs on a user's dashboard. It is an extension…☆10Feb 12, 2020Updated 6 years ago
- Rethinking Graph Regularization for Graph Neural Networks (AAAI2021)☆34Jun 6, 2021Updated 5 years ago
- Quantum tomography on optical two qubit states☆12Jul 19, 2020Updated 6 years ago
- Posted at AAAI 2023☆11Sep 4, 2025Updated 10 months ago
- PlaNet: Learning Latent Dynamics for Planning from Pixels☆10Feb 13, 2020Updated 6 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 基于Deep Qlearning Network的股票交易模型☆57May 15, 2017Updated 9 years ago
- Hybrid Computational Offloading☆16Jul 6, 2022Updated 4 years ago
- Asilomar 2020 code for Deep Actor-Critic Learning for Distributed Power Control in Wireless Mobile Networks☆42Jul 27, 2020Updated 5 years ago
- This is a official code implementation for Nonlinear RISE based Integral Reinforcement Learning algorithms for perturbed Bilateral Teleop…☆24Mar 26, 2025Updated last year
- Deep Q-Network (DQN) with Prioritized Experience Replay (PER)☆17Jan 1, 2020Updated 6 years ago
- Deep Reinforcement Learning and BCD to solve phase shift and resource allocation of RIS and RSU☆32Jan 18, 2021Updated 5 years ago
- 一个基于SSM+JSP的博客☆13Feb 27, 2022Updated 4 years ago
- ☆12Jul 15, 2020Updated 6 years ago
- Parser for files in OpenDRIVE format, offers additional functions to navigate through the road network☆12Sep 6, 2017Updated 8 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Gym Environment for AUV docking procedure☆11Sep 20, 2022Updated 3 years ago
- Not All Patches Are Equal: Hierarchical Dataset Condensation for Single Image Super-Resolution☆10May 7, 2024Updated 2 years ago
- [NeurIPS 2022] Leveraging Factored Action Spaces for Efficient Offline RL in Healthcare. https://arxiv.org/abs/2305.01738☆11Nov 27, 2022Updated 3 years ago
- Multi-view Reinforcement Learning☆11Feb 9, 2020Updated 6 years ago
- Trying to come up with an innovative robotic arm trajectory generating controller.☆17Sep 18, 2020Updated 5 years ago
- Implementation of Mutan+ArticleNet on OKVQA☆10Jan 11, 2021Updated 5 years ago
- Exploiting Inter-sample and Inter-feature Relations in Dataset Distillation (CVPR24)☆10Jun 16, 2024Updated 2 years ago
- TCP/IP server and client for Matlab☆13Mar 17, 2016Updated 10 years ago
- TensorFlow implementation of "A Relational Intervention Approach for Unsupervised Dynamics Generalization in Model-Based Reinforcement Le…☆16Jul 2, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Experiments on Model-Agnostic Meta-Learning on Few-Shot Image Classification and Meta-RL (Meta-World)☆17Mar 30, 2021Updated 5 years ago
- ☆12Jul 30, 2025Updated 11 months ago
- ☆13Jul 2, 2020Updated 6 years ago
- ☆12Mar 17, 2020Updated 6 years ago
- Implementation of DDPG+HER on gym robotics environment FetchReach-v1☆33Nov 13, 2018Updated 7 years ago
- Implementation of GAIL and AIRL using chinerrl☆16Jun 21, 2022Updated 4 years ago
- ☆14Mar 18, 2018Updated 8 years ago