Mxnet implementation of Deep Reinforcement Learning papers, such as DQN, PG, DDPG, PPO
☆28Dec 8, 2022Updated 3 years ago
Alternatives and similar repositories for Deep-rl-mxnet
Users that are interested in Deep-rl-mxnet are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- No Dependency Scala Machine Learning Algorithm Gallery☆34Sep 5, 2026Updated 2 weeks ago
- pytorch implementation of DQN, NAF, DDPG☆13Jun 7, 2018Updated 8 years ago
- Deep Reinforcement Learning with pytorch & visdom☆14May 29, 2020Updated 6 years ago
- later☆10Jul 9, 2022Updated 4 years ago
- Collection of reinforcement learning algorithms☆16Sep 29, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- PyTorch implementation of R2D2 (Recurrent Replay Distributed DPG (not DQN))☆14Mar 22, 2019Updated 7 years ago
- ☆55Jan 30, 2020Updated 6 years ago
- version 2 of the garbage openvtuber☆11Dec 7, 2022Updated 3 years ago
- Computational Graphs in R☆12Apr 16, 2020Updated 6 years ago
- ☆12Aug 22, 2011Updated 15 years ago
- A path planning framework based on Sampling-based algorithm and Deep Reinforcement learning.☆10May 9, 2023Updated 3 years ago
- ☆23Mar 14, 2020Updated 6 years ago
- ☆18Oct 4, 2024Updated last year
- Autonomous Driving on Carla simulator using Deep Deterministic Policy Gradients. Based on Kendall, et. al. 2018.☆13Apr 2, 2019Updated 7 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆12Apr 4, 2024Updated 2 years ago
- ☆13Mar 4, 2022Updated 4 years ago
- ☆16Apr 8, 2021Updated 5 years ago
- 📖 Paper: Deep Reinforcement Learning with Double Q-learning 🕹️☆62May 9, 2024Updated 2 years ago
- Setting up DDPG based reinforcement learning in ROS Gazebo environment☆14Jul 29, 2019Updated 7 years ago
- Keras Implementation of TD3(Twin Delayed DDPG) with PER(Prioritized Experience Replay) option on OpenAI gym framework☆11May 29, 2021Updated 5 years ago
- My notes while learning datascience☆10May 26, 2018Updated 8 years ago
- Implementation of a Deep Reinforcement Learning agent that is capable to share the last-level-cache of a multi-core system, between a Lat …☆10Nov 10, 2021Updated 4 years ago
- ☆12Nov 23, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- DRL-based collision avoidance for turtlebot3☆19Feb 6, 2023Updated 3 years ago
- ☆15Oct 4, 2019Updated 6 years ago
- LZ4 Data Compression and Decompression for R☆10Dec 21, 2020Updated 5 years ago
- Simulation code for the paper "Joint Resource Allocation and String-Stable CACC Design with Multi-Agent Reinforcement Learning"☆12May 17, 2023Updated 3 years ago
- ☆11Jan 18, 2022Updated 4 years ago
- Test MediaPipe Holistic JS Model☆11Feb 11, 2023Updated 3 years ago
- A local arena for Coders Strike Back (codingame.com) which tries to gauge the speed of runners in a collisionless environment.☆12May 31, 2020Updated 6 years ago
- Source code associated with final project for Machine Learning Course (CS 229) at Stanford University; Used reinforcement learning approa…☆31May 21, 2016Updated 10 years ago
- ☆12Sep 8, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- IntelliJ Idea Plugin for AI.codes☆22Feb 2, 2017Updated 9 years ago
- ☆14Apr 6, 2021Updated 5 years ago
- Implementation of Bayesian Sum-Product Networks☆13May 19, 2020Updated 6 years ago
- This repository provides a GitHub Action for running the Kani Rust Verifier in CI.☆13Updated this week
- Variational Discriminator Bottleneck: Improving Imitation Learning, Inverse RL, and GANs by Constraining Information Flow - Tensorlfow Im…☆13Feb 2, 2019Updated 7 years ago
- FlowCutter submission to PACE 2016☆12Sep 20, 2016Updated 10 years ago
- Giter8 template for Scala.js projects☆19Mar 1, 2020Updated 6 years ago