主要利用QLearning,DQN,ImprovedDQN(Ddouble DQN) 解决gym框架下的三个问题CartPole-v0,MountainCar-v0,Acrobot-v1
☆14Jan 14, 2018Updated 8 years ago
Alternatives and similar repositories for ReforceLearning
Users that are interested in ReforceLearning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- openDog visualized in ROS - forked from☆11Jun 2, 2018Updated 8 years ago
- Solving the OpenAI Gym (MountainCarContinuous-v0) with DDPG☆21Jan 23, 2023Updated 3 years ago
- The continuous mountain car problem solved with DDPG☆13Apr 19, 2020Updated 6 years ago
- Re-produce DQN, REINFORCE, REINFORCE with baseline, one-step AC, QAC, QAC with shared network, PPO2, DDPG, TD3, SAC, SAC discrete,A2C,A3C☆21Jul 27, 2020Updated 6 years ago
- ☆15Jul 25, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Code for Anxiety Level Detection from physiological signals using supervised ML algorithms☆13Dec 31, 2021Updated 4 years ago
- Control with Deep Reinforcement Learning☆16Sep 14, 2023Updated 3 years ago
- Official repository of "Minibatch optimal transport distances; analysis and applications" (https://arxiv.org/pdf/2101.01792.pdf)☆10Oct 15, 2021Updated 4 years ago
- 天池竞赛数据挖掘之二手车交易价格预测大赛☆10Mar 27, 2020Updated 6 years ago
- A synthetic approach is proposed for hand motion recognition with surface EMG signals. We used CNN features which were automatically extr…☆21Feb 21, 2019Updated 7 years ago
- ☆15Mar 30, 2020Updated 6 years ago
- Master thesis spring 2019. Template to be futher used by the department of chemical engineering at NTNU,☆39Jan 11, 2024Updated 2 years ago
- a simple test for understanding the theory of GAN, [matlab code]☆12Nov 20, 2017Updated 8 years ago
- A simple test for GAN☆10Mar 25, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆12May 19, 2021Updated 5 years ago
- C implementation of RL and IRL algorithms☆19Jul 6, 2020Updated 6 years ago
- a quaternion-based Unscented Kalman Filter on IMU to estimate quadrotor orientation. With estimates and camera data, a sphere panorama is…☆15Mar 1, 2018Updated 8 years ago
- Playing Mountain-Car without reward engineering, by combining DQN and Random Network Distillation (RND)☆41Jan 28, 2019Updated 7 years ago
- RRT*(RRT Star)-based algorithms for Path Planning of Autonomous Driving, in Python2.☆12Jun 28, 2020Updated 6 years ago
- A version of scikit-learn that includes implementations of Wager & Athey and Scott Powers causal forests.☆21Nov 9, 2016Updated 9 years ago
- ☆33Jun 5, 2026Updated 3 months ago
- Path planning & real time, image based, single camera localization on a Waveshare AlphaBot robotic platform (rpi). Code for paper "Single…☆14Jul 13, 2020Updated 6 years ago
- 图网络,深度学习,☆21Mar 10, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A ROS package for multi-robot message transport☆11Nov 27, 2024Updated last year
- HMLET (Hybrid-Method-of-Linear-and-non-linEar-collaborative-filTering-method)☆20Feb 28, 2022Updated 4 years ago
- Artificial Neural Networks (ANN), k-Nearest Neighbors, Random Forest classifier and Support Vector Machines (SVM) were trained over a HAR…☆24Mar 26, 2018Updated 8 years ago
- Extensive study and research on Udacity Self-driving Car Challenge 2☆10Dec 11, 2021Updated 4 years ago
- ☆16May 27, 2026Updated 3 months ago
- Motion Primitives based Path Planning with RRT☆12Oct 31, 2022Updated 3 years ago
- [RSS 2026] Model-Based Diffusion Optimal Control for Multi-Robot Motion Planning☆26Sep 6, 2026Updated 2 weeks ago
- ☆10Dec 10, 2021Updated 4 years ago
- Interface definitions for the Compute@Edge platform in witx.☆15Feb 11, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A docker image to interface with the MIT Racecar.☆16Feb 12, 2024Updated 2 years ago
- Mountain Car problem solving using RL - QLearning with OpenAI Gym Framework☆10Mar 20, 2018Updated 8 years ago
- Implementation of Paper "ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation"☆18Mar 23, 2026Updated 5 months ago
- ☆24Mar 13, 2018Updated 8 years ago
- Pytorch version of the MPC in model-based reinforcement learning (MBRL), currently only test in the CartPole-swing-up environment☆91Jul 25, 2020Updated 6 years ago
- Traditional feature tracking techniques such as SIFT, SURF, and Lucas Kanade algorithms define key points in terms of finding poles and c…☆12May 10, 2023Updated 3 years ago
- Low level control interface.☆16Jun 5, 2025Updated last year