强化学习大作业1 倒立摆
☆20Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for Inverted-Pendulum
Users that are interested in Inverted-Pendulum are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 利用强化学习的Q价值迭代,Q学习以及SARSA方法解决小车爬山以及倒立摆的控制问题☆15Jul 25, 2019Updated 7 years ago
- 旋转倒立摆matlab物理模型仿真☆11Jun 5, 2022Updated 4 years ago
- ☆19Sep 6, 2017Updated 8 years ago
- 4G模块接入阿里云-实现数据上传和命令下发 使用4G模块EC600S和32单片机实现接入阿里云服务器,上传光照数据和下发命令控制LED灯(PC13),同时可以打电话、发短信。详细内容看:http://t.csdn.cn/DjJCf☆18Apr 12, 2022Updated 4 years ago
- the implementation of Q_Learning☆18Jun 12, 2019Updated 7 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Achieves path following of an autonomous surface vessel (ASV) with a Deep Q-Network (DQN) agent. This code uses reinforcement learning to…☆15Jan 28, 2025Updated last year
- CFR-based Texas Hold'em AI☆11Jan 30, 2021Updated 5 years ago
- ☆35Sep 5, 2020Updated 5 years ago
- Sim2Real Transfer for Deep Reinforcement Learning with Stochastic State Transition Delays, CORL-2020.☆26Jun 3, 2021Updated 5 years ago
- Collision Avoidance simulator for USV using Deep RL. A result of TTK4550 Fordypningsoppgave at NTNU☆21Mar 21, 2024Updated 2 years ago
- ☆21Jun 12, 2025Updated last year
- Numerical Simulation of 1-D Sod Shock Tube (MATLAB Codes)☆23Aug 11, 2022Updated 3 years ago
- This branch contain the java classes for orekit-python-wrapper☆20May 13, 2026Updated 2 months ago
- 本书作者是来自日本的Yutaro Ogawa(小川熊太郎),作者的github上源码是日文注释的,这个repository把它翻译成中文☆22Dec 2, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for Towards Unifying Behavioral and Response Diversity for Open-ended Learning in Zero-sum Games☆24Feb 27, 2022Updated 4 years ago
- ☆22May 20, 2021Updated 5 years ago
- About Pytorch implementation of "DA-Ada: Learning Domain-Aware Adapter for Domain Adaptive Object Detection" (NeurIPS 2024)☆20Aug 24, 2025Updated 11 months ago
- 这个仓库用于存储一些强化学习练手小项目与算法实验。具体来讲,就是不至于单独成一个 repo 的项目,但是又值得拿出来讨论的代码。☆28May 27, 2021Updated 5 years ago
- Pytorch implementation of "Learning Domain-Aware Detection Head with Prompt Tuning" (NeurIPS 2023)☆24Mar 6, 2024Updated 2 years ago
- Domain-specific preference (DSP) data and customized RM fine-tuning.☆25Mar 7, 2024Updated 2 years ago
- Revisiting Discrete Soft Actor-Critic Accepted by Transactions on Machine Learning Research (TMLR)☆30Nov 23, 2024Updated last year
- Code for "AutoCFR: Learning to Design Counterfatual Regret Minimization Algorithms", AAAI 2022 (Oral)☆22Apr 22, 2024Updated 2 years ago
- obstacle avoidance code and algorithms.☆27Sep 2, 2020Updated 5 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Repository for "General Flow as Foundation Affordance for Scalable Robot Learning"☆70Dec 20, 2024Updated last year
- Reinforcement Learning | Multi-Agent RL | Self-Play | Proximal Policy Optimization Algorithm (PPO) agent | Unity Tennis environment☆20Dec 2, 2025Updated 7 months ago
- Reinforcement learning☆34Oct 20, 2025Updated 9 months ago
- (NeurIPS 2021) Neural Auto-Curricula in Two-Player Zero-Sum Games.☆28Nov 19, 2021Updated 4 years ago
- OpenAI gym environment for collision avoidance and path following with an AUV☆35Aug 12, 2019Updated 6 years ago
- ☆27Aug 31, 2025Updated 10 months ago
- ☆55Mar 16, 2026Updated 4 months ago
- 中国科学院大学研究生学位论文中期报告LaTex模板☆21Sep 23, 2024Updated last year
- notes☆34Jun 28, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆56Feb 6, 2021Updated 5 years ago
- This is yangxx's repo for machine learning☆36May 1, 2022Updated 4 years ago
- Instruction Following Agents with Multimodal Transforemrs☆54Nov 3, 2022Updated 3 years ago
- A simple demo for SAM+MMDetection☆51Apr 12, 2023Updated 3 years ago
- Actor Critic model to play Cartpole game☆52Aug 4, 2018Updated 7 years ago
- DRL with population coded spiking neural network for optimal and energy-efficient continuous control.☆70Dec 15, 2021Updated 4 years ago
- OpenAI gym environment of an Unmanned Surface Vehicle.☆50Apr 6, 2021Updated 5 years ago