The state-of-art deep rl algorithms for Montezuma's revenge
☆28Oct 28, 2018Updated 7 years ago
Alternatives and similar repositories for rl-montezuma
Users that are interested in rl-montezuma are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- Code for the blog post "Learning Montezuma’s Revenge from a Single Demonstration"☆207Nov 22, 2018Updated 7 years ago
- Pytorch implementation of "FeUdal Networks for Hierarchical Reinforcement Learning" for Montezuma's Revenge☆95Jul 27, 2022Updated 4 years ago
- Reusable, Easy-to-use Uncertainty module package built with Tensorflow, Keras☆14Dec 31, 2018Updated 7 years ago
- ☆39Jul 29, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Hierarchical Online Planning and Reinforcement Learning on Taxi☆32Oct 23, 2017Updated 8 years ago
- ☆10Apr 21, 2017Updated 9 years ago
- Reinforcement Learning for Classical Planning☆12Apr 13, 2022Updated 4 years ago
- Repository for slides & codes of RL Korea Bootcamp☆41Oct 28, 2019Updated 6 years ago
- IPyHOP is a Re-entrant Iterative GTPyHOP written in Python 3. PyHOP is an acronym for Python Hierarchical Ordered Planner.☆12Aug 12, 2022Updated 3 years ago
- Repository for studying distributional rl☆30Feb 2, 2025Updated last year
- weekly reinforcement learning paper reviews☆33Jan 8, 2018Updated 8 years ago
- ☆17Feb 25, 2020Updated 6 years ago
- ☆119Jul 9, 2020Updated 6 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Course Website for "AI618: Generative Model and Unsupervised Learning"☆36May 23, 2023Updated 3 years ago
- Policy Gradient algorithms (REINFORCE, NPG, TRPO, PPO)☆371Aug 1, 2019Updated 7 years ago
- ☆16Dec 8, 2022Updated 3 years ago
- ☆11Oct 3, 2022Updated 3 years ago
- ☆19Jul 23, 2025Updated last year
- [COLM 2026] An efficient 3D sampling method for long-CoT LLM.☆16May 25, 2025Updated last year
- Cochlear.ai submission for dcase2018 task2☆15Sep 14, 2018Updated 7 years ago
- Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstractions and Intrinsic Motivation☆89Mar 5, 2018Updated 8 years ago
- Neural model of hierarchical reinforcement learning☆16Sep 14, 2017Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- World Models applied to the Open AI Sonic Retro Contest☆78Jun 30, 2018Updated 8 years ago
- dqn autoplay mario bros☆21Jul 24, 2017Updated 9 years ago
- Cornell House Agent Learning Environment☆47Jun 22, 2022Updated 4 years ago
- Convolutional neural networks for sound classification☆20Dec 30, 2017Updated 8 years ago
- Official python implementation of ASGRL in ICML 2022 paper: Leveraging Approximate Symbolic Models for Reinforcement Learning via Skill D…☆20Oct 5, 2022Updated 3 years ago
- ☆14Mar 9, 2020Updated 6 years ago
- Exploring the use of options in creating small worlds for faster learning in RL Domains☆16Jan 23, 2012Updated 14 years ago
- [ICML 2023] Official code for "DevFormer: A Symmetric Transformer for Context-Aware Device Placement"☆23Dec 7, 2024Updated last year
- Random network distillation on Montezuma's Revenge and Super Mario Bros.☆55May 12, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Vue学习笔记,在线阅读地址:https://relph1119.github.io/vue-learning-notes/#/☆17Jun 19, 2023Updated 3 years ago
- Official implementation of Neural Episodic Control with State Abstraction☆13Aug 3, 2023Updated 2 years ago
- Code for the Reset-free Trial and Error learning paper (RTE) experiments☆10Jan 3, 2018Updated 8 years ago
- Some notes and code test about Deep Learning☆15Jul 12, 2020Updated 6 years ago
- Implementation of the Hierarchical and Interpretable Skill Acquisition in Multi-task Reinforcement Learning by Tianmin Shu, Caiming Xiong…☆11Jun 18, 2018Updated 8 years ago
- Fast asynchronous GPU monitoring tool across multiple machines through SSH☆12Nov 26, 2024Updated last year
- presentations☆44Dec 8, 2018Updated 7 years ago