Exploring the use of options in creating small worlds for faster learning in RL Domains
☆16Jan 23, 2012Updated 14 years ago
Alternatives and similar repositories for Small-World-RL
Users that are interested in Small-World-RL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of the Prioritized Option-Critic on the Four-Rooms Environment☆17Dec 24, 2017Updated 8 years ago
- Reinforcement Learning papers on exploration methods.☆19Jun 27, 2021Updated 5 years ago
- paper <<Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation>> python implementation☆10Mar 27, 2018Updated 8 years ago
- Reasoning about pragmatics with neural listeners and speakers☆22Apr 2, 2016Updated 10 years ago
- ☆30Jun 27, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- yet another reinforcement learning package☆12May 24, 2022Updated 4 years ago
- ☆29Mar 1, 2018Updated 8 years ago
- A tutorial on doing RL research in Julia using both Jupyter notebooks and normal project structures.☆12Jun 23, 2021Updated 5 years ago
- An implementation of the Escape Room domain for Hierarchical Reinforcement Learning.☆25May 15, 2019Updated 7 years ago
- Some notes and code test about Deep Learning☆15Jul 12, 2020Updated 6 years ago
- Jaxplorer is a Jax reinforcement learning (RL) framework for exploring new ideas.☆12Jul 19, 2024Updated 2 years ago
- iQRL: implicitly Quantized Representations for Sample-efficient Reinforcement Learning☆12Jan 8, 2025Updated last year
- An environment for tabular Reinforcement Learning agents.☆14Jun 13, 2018Updated 8 years ago
- ☆10Apr 24, 2021Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This repository contains the code base for the paper "Gait in Eight: Efficient On-Robot Learning for Omnidirectional Quadruped Locomotion…☆16Mar 8, 2025Updated last year
- ☆19Jan 30, 2023Updated 3 years ago
- Option Critic with subgoal discovery by spectral decomposition of the Successor Features Matrix or clustering in Successor features space…☆24Nov 29, 2018Updated 7 years ago
- Minimizing Control for Credit Assignment with Strong Feedback☆14Nov 3, 2024Updated last year
- ☆12Jun 8, 2020Updated 6 years ago
- Codification used for the AAMAS-17 paper "Simultaneously Learning and Advising in Multiagent Reinforcement Learning"☆15Dec 18, 2017Updated 8 years ago
- Simple implementation of the model presented in Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic …☆16Jan 22, 2019Updated 7 years ago
- An implementation of Compositional Attention: Disentangling Search and Retrieval by MILA☆14Jun 1, 2022Updated 4 years ago
- ☆18Apr 8, 2020Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Apply reinforcement learning to visual attention☆18Oct 13, 2016Updated 9 years ago
- Encrypt/decrypt files and directories using your YubiKey☆16May 3, 2026Updated 4 months ago
- Cast screen and audio on GNU/Linux to a browser using Media Source Extensions☆14Jul 9, 2026Updated last month
- A Tensorflow implementation of the Option-Critic Architecture☆75Jun 1, 2017Updated 9 years ago
- Nothing to see here, move along.☆20May 28, 2012Updated 14 years ago
- 注释版☆10Apr 29, 2017Updated 9 years ago
- Possion Reconstruction☆12Aug 9, 2021Updated 5 years ago
- Annotated bibliographies.☆40Aug 25, 2019Updated 7 years ago
- pres/v/g presentation tool for the web: present, annotate, print☆19Mar 22, 2012Updated 14 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Learning with latent language☆51Mar 28, 2021Updated 5 years ago
- Registration of 3D triangular meshes onto a 2D image can be performed using optimisation and fast X-ray simulation on GPU. Automatic esti…☆11Aug 28, 2019Updated 7 years ago
- A proxy for reverse engineering a communication protocol☆10Jan 17, 2021Updated 5 years ago
- AVPipe :-)☆12Jul 16, 2021Updated 5 years ago
- A curated list of papers presented in the 📖"Flexible Learning Reading Group" @ TU Berlin. Join us! 🤗☆27Dec 17, 2020Updated 5 years ago
- The state-of-art deep rl algorithms for Montezuma's revenge☆28Oct 28, 2018Updated 7 years ago
- Implementation of Receding Horizon Curiosity Algrithm☆13Mar 24, 2023Updated 3 years ago