强化学习大作业1 倒立摆
☆20Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for Inverted-Pendulum
Users that are interested in Inverted-Pendulum are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 利用强化学习的Q价值迭代,Q学习以及SARSA方法解决小车爬山以及倒立摆的控制问题☆15Jul 25, 2019Updated 7 years ago
- ☆19Sep 6, 2017Updated 8 years ago
- 4G模块接入阿里云-实现数据上传和命令下发 使用4G模块EC600S和32单片机实现接入阿里云服务器,上传光照数据和下发命令控制LED灯(PC13),同 时可以打电话、发短信。详细内容看:http://t.csdn.cn/DjJCf☆18Apr 12, 2022Updated 4 years ago
- Hybrid Computational Offloading☆16Jul 6, 2022Updated 4 years ago
- the implementation of Q_Learning☆18Jun 12, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- CFR-based Texas Hold'em AI☆11Jan 30, 2021Updated 5 years ago
- 酒店管理系统,利用QT工具实现,图形化界面提供客房、员工管理所需的操作☆11Feb 20, 2020Updated 6 years ago
- ☆35Sep 5, 2020Updated 5 years ago
- Collision Avoidance simulator for USV using Deep RL. A result of TTK4550 Fordypningsoppgave at NTNU☆21Mar 21, 2024Updated 2 years ago
- ☆15Oct 6, 2019Updated 6 years ago
- The reinforcement learning training code for AgiBot X1.☆15Jan 15, 2025Updated last year
- Code for Towards Unifying Behavioral and Response Diversity for Open-ended Learning in Zero-sum Games☆24Feb 27, 2022Updated 4 years ago
- ☆22May 20, 2021Updated 5 years ago
- 这个仓库用于存储一些强化学习练手小项目与算法实验。具体来讲,就是不至于单独成一个 repo 的项目,但是又值得拿出来讨论的代码。☆28May 27, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Revisiting Discrete Soft Actor-Critic Accepted by Transactions on Machine Learning Research (TMLR)☆30Nov 23, 2024Updated last year
- Repository for "General Flow as Foundation Affordance for Scalable Robot Learning"☆70Dec 20, 2024Updated last year
- Reinforcement Learning | Multi-Agent RL | Self-Play | Proximal Policy Optimization Algorithm (PPO) agent | Unity Tennis environment☆20Dec 2, 2025Updated 9 months ago
- Reinforcement learning☆34Oct 20, 2025Updated 10 months ago
- (NeurIPS 2021) Neural Auto-Curricula in Two-Player Zero-Sum Games.☆28Nov 19, 2021Updated 4 years ago
- OpenAI gym environment for collision avoidance and path following with an AUV☆36Aug 12, 2019Updated 7 years ago
- Team SINGABOAT-VRX's GitHub Repository for Virtual RobotX (VRX) Competition.☆56Nov 6, 2022Updated 3 years ago
- notes☆34Jun 28, 2022Updated 4 years ago
- Deep RL Code for XDO: A Double Oracle Algorithm for Extensive-Form Games☆40Aug 27, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Actor Critic model to play Cartpole game☆52Aug 4, 2018Updated 8 years ago
- OpenAI gym environment of an Unmanned Surface Vehicle.☆52Apr 6, 2021Updated 5 years ago
- An official implementation of Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards☆36Oct 3, 2025Updated 11 months ago
- Master thesis repo for implementation of Model Predictive Controller and reinforcement learning (RL) controller☆51Jan 16, 2024Updated 2 years ago
- 动手学强化学习代码☆67Jan 17, 2024Updated 2 years ago
- Official Code Release for Pipeline PSRO: A Scalable Approach for Finding Approximate Nash Equilibria in Large Games☆58Aug 30, 2024Updated 2 years ago
- Isaac Gym Environments for Legged Robots☆53Oct 1, 2025Updated 11 months ago
- The implementation of LSTM-TD3.☆87Feb 14, 2023Updated 3 years ago
- USV simulator for ROS Melodic and Gazebo 9. Copied from an old version of https://github.com/osrf/vrx☆81Nov 5, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for reproducing results in GraphMix paper☆72Nov 22, 2022Updated 3 years ago
- Segment Policy Optimization: Effective Segment-Level Credit Assignment in RL for Large Language Models☆56Sep 19, 2025Updated 11 months ago
- ☆165Dec 1, 2021Updated 4 years ago
- Mamba-YOLO-World: Marrying YOLO-World with Mamba for Open-Vocabulary Detection☆104Mar 12, 2025Updated last year
- Highway driving simulator incorporating NGSIM dataset using reinforcement learning☆83Apr 22, 2022Updated 4 years ago
- The official code releasement of publications in MARL field of TJU RL lab.☆89Jul 15, 2022Updated 4 years ago
- GelSight SDK for robotic sensors☆197Jun 25, 2025Updated last year