My own implementation of Reinforcement Learning algorithms using Tensorflow 2.0
☆30Jan 22, 2022Updated 4 years ago
Alternatives and similar repositories for rl-tf2
Users that are interested in rl-tf2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- End-to-end reinforcement learning using DDPG and PPO algorithms in a simulated robot environment☆21Mar 13, 2020Updated 6 years ago
- Official PyTorch code for "Recurrent Off-policy Baselines for Memory-based Continuous Control" (DeepRL Workshop, NeurIPS 21)☆95Nov 21, 2023Updated 2 years ago
- using recurrent networks(LSTM) to solve POMDPs☆35Oct 10, 2018Updated 7 years ago
- RDFS: an erasure code based cloud storage system☆39Jul 28, 2014Updated 12 years ago
- The implementation of LSTM-TD3.☆87Feb 14, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation of the Discrete Soft Actor-Critic algorithm with RNN policy in PyTorch☆26Jan 7, 2023Updated 3 years ago
- Training Agents in a cooperative multi-agent deep reinforcement learning setting to transport objects across a space☆14Jul 5, 2021Updated 5 years ago
- SeMoDe is a tool to support lifecycle activites of Serverless functions on different platforms. Currently automated test generation on AW…☆14Jul 26, 2023Updated 3 years ago
- later☆10Jul 9, 2022Updated 4 years ago
- Repository for the paper "Discovering and Categorising Language Biases in Reddit" accepted at the International Conference on Web and Soc…☆12Aug 20, 2024Updated 2 years ago
- Multi objective optimization-based routing algorithm for SDN networks☆20Aug 31, 2020Updated 6 years ago
- Python code for implementation of the paper 'Reinforcement Learning-Based Adaptive PID Controller for DPS☆16Aug 28, 2020Updated 6 years ago
- Hierarchical and Stable Multiagent Reinforcement Learning for Cooperative Navigation Control☆14May 5, 2022Updated 4 years ago
- This is a project based on machine learning and deep learning method for playing Gobang by controlling mechanical arm(利用机械臂下五子棋)☆13Apr 16, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This repository provides the python implementation for the paper "Decentralized Multi-Agent Formation Control via Deep Reinforcement Lear…☆20Jan 19, 2022Updated 4 years ago
- Application of an LSTM-based policy gradient on an RL agent☆14Aug 24, 2022Updated 4 years ago
- PyTorch Implementation of the RDPG (Recurrent Deterministic Policy Gradient)☆55Dec 8, 2022Updated 3 years ago
- Keras Implementation of DDPG(Deep Deterministic Policy Gradient) with PER(Prioritized Experience Replay) option on OpenAI gym framework☆13Mar 25, 2023Updated 3 years ago
- A path planning framework based on Sampling-based algorithm and Deep Reinforcement learning.☆10May 9, 2023Updated 3 years ago
- 基于深度强化学习不同算法的移动机器人导航避障☆20Jul 6, 2021Updated 5 years ago
- ☆20Sep 14, 2019Updated 6 years ago
- NSMC Satellite Product Data Reader (AWX)☆21Aug 16, 2024Updated 2 years ago
- Autonomous visual navigation using the depth images☆11Aug 15, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Long-distance maritime polar route planning, taking into account complex changing environmental conditions.☆24Updated this week
- Implementation of Population-Guided Parallel Policy Search for Reinforcement Learning☆22Jan 9, 2020Updated 6 years ago
- This is a DQN-based recommendation system for item-list recommendation and it finally achieved second place in the competition of RL-base…☆11Oct 8, 2021Updated 4 years ago
- 自制的爱斯维尔期刊的LaTeX投稿工作流程模板☆23Sep 19, 2024Updated last year
- Keras Implementation of TD3(Twin Delayed DDPG) with PER(Prioritized Experience Replay) option on OpenAI gym framework☆11May 29, 2021Updated 5 years ago
- Cuda-Engined Adaptive Optics☆28Oct 21, 2025Updated 10 months ago
- The PPO algorithm based on the route planning of the ship's path at the sea☆20Jul 5, 2023Updated 3 years ago
- ☆21Nov 16, 2022Updated 3 years ago
- ☆10Sep 21, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of Deep Deterministic Policy Gradient (DDPG) with Prioritized Experience Replay (PER)☆54Apr 29, 2026Updated 4 months ago
- Hide the memory of the process in the Linux kernel.☆10Dec 8, 2020Updated 5 years ago
- ☆14Feb 6, 2025Updated last year
- ☆19May 12, 2021Updated 5 years ago
- Official Code of Decoupled Graph Convolution (DGC)☆16Jan 31, 2026Updated 7 months ago
- ICLR Reproducibility Challenge for Discriminator-Actor-Critic☆20Jan 7, 2019Updated 7 years ago
- ☆21Oct 2, 2016Updated 9 years ago