Open AI Gym - Pendulum-v1 reinforcement learning (DQN, SAC)
☆21Jan 26, 2024Updated 2 years ago
Alternatives and similar repositories for rl-pendulum
Users that are interested in rl-pendulum are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MATLAB implementation of DQN for a navigation environment☆13Aug 13, 2020Updated 6 years ago
- This is the code implementation of the Neural ordinary differential equations-based Lyapunov-Barrier Actor-Critic (NLBAC)☆17Sep 4, 2024Updated last year
- Training an autonomous driving robot in Gazebo Simulator by soft actor critic method☆15Apr 12, 2022Updated 4 years ago
- Crypto-Options Volatility Surface Calibration and Arbitrage☆17Dec 26, 2022Updated 3 years ago
- Personalized Client-Edge-Cloud Hierarchical Federated Learning on Non-IID Data☆11Sep 7, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Working examples of Deep Q Network of Reinforcement Learning☆14Apr 7, 2020Updated 6 years ago
- Implementation of Quantile-Constrained Policy Optimization (QCPO)☆11Sep 28, 2022Updated 3 years ago
- Implementation of Pareto Deep Q Networks in a multi-objective Gym Reinforcement Learning Environment☆18Jun 19, 2023Updated 3 years ago
- ☆17Oct 25, 2023Updated 2 years ago
- python script to bridge between a mqtt server (e.g. mosquitto) and a socketcan device☆13Jun 12, 2016Updated 10 years ago
- The PLCnext-ROS-bridge enables the whole power of the open source Roboter Operating System (ROS) for the IEC61131 world.☆11Aug 31, 2023Updated 2 years ago
- PyTorch implementation of PtrNet to solve sorting problem.☆12Dec 19, 2017Updated 8 years ago
- The aim of this repo is to bring ideas and relevant literature relating to Safe-RL in the context of autonomous vehicles.☆51Aug 3, 2018Updated 8 years ago
- ☆17Sep 11, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- UAV-based path planning for efficient localization of non-uniformly distributed weeds using prior knowledge: A reinforcement-learning app…☆15Jul 1, 2025Updated last year
- ☆13Jun 26, 2020Updated 6 years ago
- ☆20Nov 21, 2023Updated 2 years ago
- Using DDPG agent to control UAV system with energy efficiency☆16Jan 7, 2023Updated 3 years ago
- ☆10Sep 7, 2017Updated 8 years ago
- ☆16Jul 25, 2023Updated 3 years ago
- [AAMAS 2025] Privacy-preserving and Personalized RLHF, with convergence guarantees. The Code contains experiments for training multiple i…☆16Apr 16, 2025Updated last year
- Code for paper "Multi-Agent Active Search: a Reinforcement Learning Approach", submitted to ICRA 2022.☆13Sep 19, 2021Updated 4 years ago
- Encode-attend-navigate unofficial Pytorch implementation☆12Oct 1, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- This project applies Monte Carlo Tree Search (MCTS) to a simple grid world.☆10May 30, 2018Updated 8 years ago
- ☆45Jan 26, 2024Updated 2 years ago
- ☆13Nov 29, 2020Updated 5 years ago
- This repository contains the implementation of a Deep Deterministic Policy Gradient (DDPG) algorithm applied to solve the Reacher environ…☆12Apr 8, 2023Updated 3 years ago
- ☆23May 13, 2021Updated 5 years ago
- The implementation of STAR-HiT.☆11Oct 18, 2023Updated 2 years ago
- Academic Study of A Multi-Agent Quadrotors (Drones) Simulator with Obstacles and Goals Using the Artificial Potential Field Approach(APF)…☆19Feb 13, 2022Updated 4 years ago
- The Paradox of Choice: Using Attention in Hierarchical Reinforcement Learning☆11Oct 31, 2021Updated 4 years ago
- some resources I collected☆13Apr 28, 2019Updated 7 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆16Dec 5, 2024Updated last year
- Analytics for Trading with NOA☆25Jan 24, 2022Updated 4 years ago
- ☆13Jun 1, 2020Updated 6 years ago
- OpenAI's Gym Car-Racing-V0 environment was tackled and, subsequently, solved using a variety of Reinforcement Learning methods including …☆23Aug 7, 2022Updated 4 years ago
- Transfer learning in deep reinforcement learning for continuous control. Implemented DDPG and TD3 algorithms and evaluated ability to ada…☆18Feb 25, 2025Updated last year
- Hierarchical and Stable Multiagent Reinforcement Learning for Cooperative Navigation Control☆14May 5, 2022Updated 4 years ago
- This repository considers the implementation of the paper "FoX: Formation-aware exploration in multi-agent reinforcement learning" which …☆26Oct 24, 2024Updated last year