Version 3.0.0 Pytorch implementations of DQN, DDQN, DDPG, SAC, Discrete SAC. With more features :)
☆12Feb 16, 2023Updated 3 years ago
Alternatives and similar repositories for RL-v2
Users that are interested in RL-v2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- Economics of Ransomware | Dataset☆15May 2, 2018Updated 8 years ago
- Training and testing pipeline for ransomware classification based on screenshots of the splash screens or ransom notes (https://arxiv.org…☆11Jul 19, 2020Updated 6 years ago
- 1. Simulation of a job shop production system 2. Reinforcement Learning agent to control the production system☆11Sep 8, 2021Updated 4 years ago
- Implementation for mSAC methods in PyTorch☆42Oct 10, 2021Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆10Sep 9, 2022Updated 3 years ago
- My Master Thesis at the ASL supervised by Hermann Blum, Francesco Milano and Dr. Cadena Cesar☆11Dec 20, 2022Updated 3 years ago
- Simulation of car parking in different parking lots using Unity ML-Agents☆13Dec 16, 2023Updated 2 years ago
- Visual Localization with an image sequence. 3DV Project @ ETH Zurich, 2022.☆20Jul 20, 2022Updated 4 years ago
- 量化交易网站,软工三大作业迭代三,团队项目☆11Mar 8, 2018Updated 8 years ago
- Simulation Architectures for Reinforcement Learning applied to Robotics☆13Sep 4, 2024Updated 2 years ago
- ☆13Oct 5, 2021Updated 4 years ago
- DQN related algorithms☆10Mar 5, 2023Updated 3 years ago
- Implementation of Multi-Agent Reinforcement Learning algorithm(s). Currently includes: MADDPG☆68Jul 9, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Prioritized Sequence Experience Replay☆10Aug 16, 2021Updated 5 years ago
- The implementation of Discriminator Soft Actor Critic☆15Jan 25, 2020Updated 6 years ago
- Demonstrating the usage of FGYM: A Toolkit for benchmarking FPGA-accelerated Reinforcement Learning☆14Aug 12, 2021Updated 5 years ago
- Man in the middle attack demo☆11Jan 14, 2018Updated 8 years ago
- [NeurIPS 2020] "FracTrain: Fractionally Squeezing Bit Savings Both Temporally and Spatially for Efficient DNN Training" by Yonggan Fu, Ha…☆10Feb 13, 2022Updated 4 years ago
- Factored Interactive POMDP solver based on symbolic Perseus.☆11Aug 12, 2025Updated last year
- A small reinforcement learning library for my masters dissertation project☆16May 9, 2026Updated 3 months ago
- ☆14Jul 27, 2022Updated 4 years ago
- ppo+action mask for atari tennis agent☆12Mar 2, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implementation of Deep Deterministic Policy Gradient (DDPG) with Prioritized Experience Replay (PER)☆54Apr 29, 2026Updated 4 months ago
- 📖 UI/UX context detection engine☆12Jan 3, 2021Updated 5 years ago
- Related papers for offline reforcement learning (we mainly focus on representation and sequence modeling and conventional offline RL)☆19Apr 21, 2022Updated 4 years ago
- Official codebase for Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings.☆21Mar 5, 2021Updated 5 years ago
- This repository contains a jupyter notebook which includes the code for Q-learning and SARSA based path planning of 2 UAVs in 2D grid env…☆11Aug 11, 2023Updated 3 years ago
- optimal power flow based on DistFlow☆19Apr 12, 2023Updated 3 years ago
- ☆17Jun 7, 2017Updated 9 years ago
- A Linux/Windows Ransomware PoC written in Python, Go and C☆18Jun 17, 2023Updated 3 years ago
- This repo is created to perform I/O Request Packet (IRP) driven ransomware analysis where the IRP logs were collected during ransomware e…☆11Aug 14, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- path planning using Q learning algorithm☆12Oct 12, 2023Updated 2 years ago
- ☆15Feb 28, 2020Updated 6 years ago
- Sharing the codebase and steps for artifact evaluation for ISCA 2023 paper☆16Feb 20, 2024Updated 2 years ago
- ☆13Apr 29, 2021Updated 5 years ago
- ☆18Jul 25, 2024Updated 2 years ago
- Code that accompanies the PyData New York (2022) talk: Addressing the sensitivity of Large language models☆13Nov 7, 2022Updated 3 years ago
- ☆13Jun 19, 2018Updated 8 years ago