PyTorch implementation of Soft Actor-Critic (SAC), Twin Delayed DDPG (TD3), Actor-Critic (AC/A2C), Proximal Policy Optimization (PPO), QT-Opt, PointNet..
☆16Nov 18, 2020Updated 5 years ago
Alternatives and similar repositories for SOTA-RL-Algorithms
Users that are interested in SOTA-RL-Algorithms are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Deep Reinforcement Learning algorithms for Policy Value methods written from scratch.☆22Aug 27, 2020Updated 5 years ago
- [NeurIPS 2024] Maximum Entropy Reinforcement Learning via Energy-Based Normalizing Flow☆44Jun 15, 2026Updated last month
- Jaxplorer is a Jax reinforcement learning (RL) framework for exploring new ideas.☆13Jul 19, 2024Updated 2 years ago
- Implementation of Continuous Control RL Algorithms☆11Dec 8, 2022Updated 3 years ago
- Codes used to perform the experiments described in this work: https://arxiv.org/abs/1904.05803☆12Aug 29, 2019Updated 6 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- 한국산업기술대학교 카카오톡 챗봇 산돌이☆10Oct 1, 2021Updated 4 years ago
- Designing an optimized path for multiple robots in a warehouse for picking and delivery operations using A* algorithm (shortest path) and…☆11Jul 28, 2023Updated 2 years ago
- 문장단위로 분절된 나무위키 데이터셋. Releases에서 다운로드 받거나, tfds-korean을 통해 다운로드 받으세요.☆19Jun 16, 2021Updated 5 years ago
- Implementation of some of the Deep Distributional Reinforcement Learning Algorithms.☆26Jun 17, 2025Updated last year
- Aerial Combat environment build around PyFlyt☆12Aug 12, 2023Updated 2 years ago
- OpenAI Gym Environment for Puyo Puyo☆17Apr 24, 2024Updated 2 years ago
- Code for paper "Learning to Guide: Guidance Law Based on Deep Meta-learning and Model Predictive Path Integral Control"☆11May 26, 2019Updated 7 years ago
- 학교 외박신청 자동화 App Based in React Native☆13Mar 2, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆15Updated this week
- ☆11Sep 29, 2022Updated 3 years ago
- The Ranking Cost algorithm for multi-path routing of gridworld.(多智能体路径规划,电路规划)☆19Dec 20, 2021Updated 4 years ago
- 비속어 탐지 모델☆16Dec 19, 2019Updated 6 years ago
- soft q learning and soft actor critic☆16Dec 23, 2018Updated 7 years ago
- ☆10Dec 30, 2021Updated 4 years ago
- ☆22Nov 17, 2020Updated 5 years ago
- Minimum Energy Resource Allocation Strategy with partial offloading☆10Jan 17, 2022Updated 4 years ago
- Jupyter notebook with the code of a probabilistic neural network in PyTorch☆12Jan 17, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- fast api with machine learning☆11Apr 23, 2023Updated 3 years ago
- Codes for blog posting☆19Nov 23, 2024Updated last year
- A minimal and interpretable Brian2 based DYNAP neuromorphic processor simulator for educational purposes.☆12Jun 23, 2022Updated 4 years ago
- ☆17Mar 31, 2022Updated 4 years ago
- ☆19Oct 12, 2022Updated 3 years ago
- ☆12Oct 29, 2022Updated 3 years ago
- Reinforcement Learning material☆24May 10, 2020Updated 6 years ago
- Official code of Nash-DQN for paper: Nash-DQN algorithm for two-player zero-sum Markov games, details see our paper: A Deep Reinforcement…☆22Aug 26, 2022Updated 3 years ago
- ♪ Programming music theory concepts☆16Mar 7, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆11Jun 15, 2019Updated 7 years ago
- Source code for paper "Trajectory of Alternating Direction Method of Multipliers and Adaptive Acceleration" of NeurIPS 2019☆10Jan 25, 2024Updated 2 years ago
- ETL (Extract, Transform and Load) with the Spark Python API (PySpark) and Hadoop Distributed File System (HDFS)☆17Dec 18, 2018Updated 7 years ago
- Library for creating smooth cubic splines☆10Oct 15, 2020Updated 5 years ago
- Order Fulfillment by Multi-Agent Reinforcement Learning☆28Jun 12, 2026Updated last month
- Solve Multi-agent Path Finding problem for heterogeneous robots.☆30Feb 24, 2021Updated 5 years ago
- Topology Aware Task Mapping Tool☆14Jul 27, 2016Updated 9 years ago