πReinforcement Learning: Super Mario Bros with dueling dqnπ
β150May 20, 2025Updated last year
Alternatives and similar repositories for Super-Mario-RL
Users that are interested in Super-Mario-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Super Mario Bros training with Ray RLlib DQN algorithmβ24May 22, 2021Updated 5 years ago
- Simple implementations of multi-agent evolutionary strategies using pytorch.β17Jan 15, 2022Updated 4 years ago
- A Deep Q Network used for running experiments on reinforcement learning agents targeted at learning Super Mario Bros (NES)β11Oct 12, 2017Updated 8 years ago
- A modular implementation for Proximal Policy Optimization in Tensorflow 2 using Eagerly Execution for the Super Mario Bros enviroment.β21Nov 6, 2019Updated 6 years ago
- An OpenAI Gym interface to Super Mario Bros. & Super Mario Bros. 2 (Lost Levels) on The NESβ872Jun 10, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β12Aug 24, 2023Updated 2 years ago
- Proximal Policy Optimization (PPO) algorithm for Super Mario Brosβ1,298Jul 24, 2021Updated 5 years ago
- An attempt at recreating DeepMind's implementation of Deep Q Learning on Atari Breakout using PyTorchβ13Jan 16, 2020Updated 6 years ago
- λ μ΄νΌ κ΅μν μ΅νμ μκ°μ¬νβ10Apr 19, 2017Updated 9 years ago
- Interactive tutorial to build a learning Mario, for first-time RL learnersβ255Jan 27, 2023Updated 3 years ago
- Natural Language Processing toolsβ12Jan 26, 2017Updated 9 years ago
- This is pytorch implmentation project of Bootsrapped DQNβ13Dec 6, 2020Updated 5 years ago
- Reinforcement Learning Based Collision Avoidance with Adaptive Environment Modeling for Crowded Scenesβ42Feb 23, 2024Updated 2 years ago
- Negative Update Intervals in Multi-Agent Deep Reinforcement Learningβ35May 14, 2019Updated 7 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implementation of Relational Deep Reinforcement Learningβ26Jan 31, 2020Updated 6 years ago
- Contextual Bandits Action Elimination DQNβ21Jun 25, 2018Updated 8 years ago
- Implementation of Bootstrap DQN and Randomized Prior Functions on ALEβ56Mar 12, 2025Updated last year
- Webots visual tracking example with OpenCVβ12Jan 17, 2023Updated 3 years ago
- Continual RL with wold models (Collas 2023)β23Dec 10, 2023Updated 2 years ago
- Tensorflow implementation of DQN to control cart-pole from OpenAI gym environmentβ14Sep 24, 2017Updated 8 years ago
- An implementation of the Augmented Random Search algorithmβ14Jan 29, 2022Updated 4 years ago
- Reinforcement Learning Robot avoiding obstacles(Python + V_rep)β12Oct 29, 2019Updated 6 years ago
- Hierarchical Deep Reinforcement Learning for Adaptive Resource Management in Integrated Terrestrial and Non-Terrestrial Networksβ17Feb 3, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- BlindTool β A mobile app that gives a "sense of vision" to the blind with deep learningβ13Sep 1, 2022Updated 3 years ago
- Interview Diaries is a user-friendly blogging platform designed for developers to effortlessly share their interview experiences.β11Apr 25, 2024Updated 2 years ago
- Official code repository for the paper "Rethinking Model Prototyping through the MedMNIST+ Dataset Collection" @ Scientific Reportsβ13Mar 5, 2025Updated last year
- Produce intelligence by means of natural selection without objective/reward optimizationβ16Sep 29, 2021Updated 4 years ago
- OpenAI Gym's LunarLander-v2 Implementationβ42Apr 27, 2024Updated 2 years ago
- code of the paper "Reliability modeling and statistical analysis of accelerated degradation process with memory effects and unit-to-unit β¦β13Jan 28, 2026Updated 6 months ago
- Gym wrapper for pysc2β10Sep 16, 2022Updated 3 years ago
- Deep Q Network-based scheduling over two-hop wireless channels.β17Apr 3, 2021Updated 5 years ago
- a script help you to auto play bard music.β16Sep 5, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β12Jun 30, 2022Updated 4 years ago
- An Gymnasium environment for the Flappy Bird gameβ113Jan 28, 2026Updated 6 months ago
- β11Oct 29, 2024Updated last year
- Simple Grid Environment for Gymnasiumβ65Mar 1, 2026Updated 5 months ago
- This project enhances the LLaMA-2 model using Quantized Low-Rank Adaptation (QLoRA) and other parameter-efficient fine-tuning techniques β¦β13Apr 18, 2024Updated 2 years ago
- VIDIMU-TOOLS is a code repository related to the public dataset "VIDIMU. multimodal video and IMU kinematic dataset on daily life activitβ¦β12Jun 2, 2024Updated 2 years ago
- Reinforcement learning tutorialsβ412Mar 25, 2023Updated 3 years ago