A Deep Q Network used for running experiments on reinforcement learning agents targeted at learning Super Mario Bros (NES)
☆11Oct 12, 2017Updated 8 years ago
Alternatives and similar repositories for Super-Mario-Bros-DQN
Users that are interested in Super-Mario-Bros-DQN are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An implementation of (Double/Dueling) Deep-Q Learning to play Super Mario Bros.☆75Jul 3, 2026Updated 2 months ago
- Reinforcement Learning Tutorial on Super Mario☆90Nov 13, 2017Updated 8 years ago
- ☆15Jul 25, 2019Updated 7 years ago
- PyTorch Implementation of DQN and training Super Mario Bros☆26Nov 16, 2025Updated 9 months ago
- deep-reinforcement-learning☆16Mar 16, 2019Updated 7 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Storm for almost everyone☆16Aug 3, 2026Updated last month
- PyTorch implementation of the End-to-End Memory Network with attention layer vizualisation support.☆12Jun 30, 2018Updated 8 years ago
- NeurIPS 2019: DQN(λ) = Deep Q-Network + λ-returns.☆25May 20, 2024Updated 2 years ago
- Cursor CLI + Claude Code + Codex Sandbox☆22Aug 19, 2026Updated 2 weeks ago
- Baum-Welch for all kind of Markov models☆24May 25, 2026Updated 3 months ago
- ☆12Jun 17, 2022Updated 4 years ago
- A minimal implementation of a VAE with BinConcrete (relaxed Bernoulli) latent distribution in TensorFlow.☆22Feb 1, 2020Updated 6 years ago
- ☆12Feb 21, 2022Updated 4 years ago
- ☆24Jul 8, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- a collection of DRL-repo in Github☆15Oct 21, 2020Updated 5 years ago
- ☆10Aug 29, 2017Updated 9 years ago
- Collection of tutorials, exercises and papers on RL☆17Oct 16, 2017Updated 8 years ago
- Estimate the fundamental frequency and inharmonicity coefficient of an isolated piano note☆11Jan 1, 2018Updated 8 years ago
- ☆69Nov 30, 2018Updated 7 years ago
- ☆19Mar 5, 2019Updated 7 years ago
- 2D bullet hell shoot'em-up☆10Sep 6, 2025Updated last year
- A full example report☆11Jul 23, 2019Updated 7 years ago
- Perfect Pitch for Anyone. Experimental visualization of music on piano keys via FFT.☆17Dec 9, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Training Agents in a cooperative multi-agent deep reinforcement learning setting to transport objects across a space☆14Jul 5, 2021Updated 5 years ago
- Reinforcement Learning PPO Super Mario Bros Agent☆13Dec 11, 2022Updated 3 years ago
- Command-line tool to manage CPython Misc/NEWS.d entries☆21Updated this week
- Deep Successor Representation☆18Mar 6, 2018Updated 8 years ago
- Rectifying Self Organizing Map☆29Oct 7, 2024Updated last year
- Implementation of Deep Q-learning from Demonstrations using Keras and a Retro Gym environment.☆14Jul 16, 2018Updated 8 years ago
- A multi-agent reinforcement learning solution to Flatland3 challenge.☆18Feb 16, 2024Updated 2 years ago
- ☆49Updated this week
- A remake of the 1990's classic mario hit. Play the game here:☆11Jun 17, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Repository hosting the KLFitter library – the Kinematic Likelihood Fitter.☆10Mar 4, 2026Updated 6 months ago
- ☆28Jul 28, 2022Updated 4 years ago
- ForgER algorithm☆23Oct 3, 2022Updated 3 years ago
- OCaml for web programming☆53Aug 28, 2016Updated 10 years ago
- This is the repository that introduces research topics related to protecting intellectual property (IP) of AI from a data-centric perspec…