Mxnet implementation of Deep Reinforcement Learning papers, such as DQN, PG, DDPG, PPO
☆28Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for Deep-rl-mxnet
Users that are interested in Deep-rl-mxnet are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [TNNLS] PGDQN: A generalized and efficient preference-guided epsilon-greedy policy equipped DQN for Atari and Autonomous Driving☆11Oct 9, 2023Updated 2 years ago
- Automated neural architecture search algorithms implemented in PyTorch and Autogluon toolkit.☆12Apr 17, 2020Updated 6 years ago
- Source code for ICLR 2024 paper "GRAPH-CONSTRAINED DIFFUSION FOR END-TO-END PATH PLANNING"☆14Jun 4, 2024Updated 2 years ago
- Reinforcement Learning completely written in C# Unity Asset☆20May 19, 2024Updated 2 years ago
- later☆10Jul 9, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Collection of reinforcement learning algorithms☆16Sep 29, 2025Updated 9 months ago
- PyTorch implementation of R2D2 (Recurrent Replay Distributed DPG (not DQN))☆14Mar 22, 2019Updated 7 years ago
- This project features a dynamic combat and traversal system inspired by Sekiro, incorporating fluid movement, precise timing, and strateg…☆14Oct 22, 2024Updated last year
- Deep Q-Network (DQN) with Prioritized Experience Replay (PER)☆17Jan 1, 2020Updated 6 years ago
- ☆18Oct 4, 2024Updated last year
- Autonomous Driving on Carla simulator using Deep Deterministic Policy Gradients. Based on Kendall, et. al. 2018.☆13Apr 2, 2019Updated 7 years ago
- 📖 Paper: Deep Reinforcement Learning with Double Q-learning 🕹️☆62May 9, 2024Updated 2 years ago
- Setting up DDPG based reinforcement learning in ROS Gazebo environment☆14Jul 29, 2019Updated 6 years ago
- Keras Implementation of TD3(Twin Delayed DDPG) with PER(Prioritized Experience Replay) option on OpenAI gym framework☆11May 29, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆12Mar 6, 2020Updated 6 years ago
- Implementation of a Deep Reinforcement Learning agent that is capable to share the last-level-cache of a multi-core system, between a Lat…☆10Nov 10, 2021Updated 4 years ago
- ☆12Nov 23, 2021Updated 4 years ago
- Mobile manipulator Task and Motion Planning(TAMP) implementaion by using legacy Method (BasePlacement). This repository is Tested in C…☆19Oct 23, 2024Updated last year
- Dead simple X11 color picker. No extra libs required.☆11May 8, 2018Updated 8 years ago
- Simulation code for the paper "Joint Resource Allocation and String-Stable CACC Design with Multi-Agent Reinforcement Learning"☆12May 17, 2023Updated 3 years ago
- Schedule for ArtOfSAT☆11Oct 11, 2023Updated 2 years ago
- ☆11Jan 18, 2022Updated 4 years ago
- Source code associated with final project for Machine Learning Course (CS 229) at Stanford University; Used reinforcement learning approa…☆31May 21, 2016Updated 10 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- An open source video conferencing tool for the XO laptop☆16Sep 20, 2013Updated 12 years ago
- A gym game for Contra that for reinforcement learning☆10Oct 18, 2021Updated 4 years ago
- Web application interface for Mathematica☆11Mar 28, 2015Updated 11 years ago
- code with the RA-L'23 paper - "Mixed Integer Programming for Time-Optimal Multi-Robot Coverage Path Planning with Heuristics"☆37Dec 11, 2025Updated 7 months ago
- ☆12Sep 29, 2021Updated 4 years ago
- ROS packages for building wide intelligence project, University of Texas at Austin☆10Jul 5, 2024Updated 2 years ago
- Basic reinforcement learning algorithms. Including:DQN,Double DQN, Dueling DQN, SARSA, REINFORCE, baseline-REINFORCE, Actor-Critic,DDPG,D…☆97Mar 1, 2021Updated 5 years ago
- ☆14Apr 6, 2021Updated 5 years ago
- Implementation of Bayesian Sum-Product Networks☆13May 19, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Algorithmic trading bot in Rust with multi-agent architecture, 10 strategies, risk management, and native egui UI. Supports Alpaca & Bina…☆15Jun 13, 2026Updated last month
- Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow - Tensorlfow Im…☆13Feb 2, 2019Updated 7 years ago
- FlowCutter submission to PACE 2016☆12Sep 20, 2016Updated 9 years ago
- Reinforcement learning of driving a racing car in TORCS using DDPG algorithm☆14Mar 3, 2018Updated 8 years ago
- transparent and reproducible analysis of merging behavior: evidence from exiD dataset☆18Apr 24, 2023Updated 3 years ago
- ☆12Jun 22, 2023Updated 3 years ago
- ☆13Nov 4, 2023Updated 2 years ago