Simple bit flipping with sparse rewards using HER, similarly to the original paper
☆39Feb 25, 2019Updated 7 years ago
Alternatives and similar repositories for Hindsight-Experience-Replay---Bit-Flipping
Users that are interested in Hindsight-Experience-Replay---Bit-Flipping are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 2D Gridworld navigation using RL with Hindsight Experience Replay☆47May 29, 2019Updated 7 years ago
- ☆13Apr 4, 2023Updated 3 years ago
- rlcourse-march-17-hugobb created by GitHub Classroom☆15Jul 3, 2024Updated 2 years ago
- Logarithmic Reinforcement Learning☆28Apr 7, 2023Updated 3 years ago
- Primary Recommender System: online[matching|ranking...](Flask|Vue) - nearline[model serving|real-time service](Flink|tensorflow serving|r…☆14Mar 27, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Dec 5, 2020Updated 5 years ago
- ☆13Dec 28, 2023Updated 2 years ago
- Comp 781 Project☆10Jan 2, 2026Updated 8 months ago
- ☆13Feb 3, 2021Updated 5 years ago
- Implementation based on the paper "Ant Colony System for Graph Coloring Problem"☆11Apr 14, 2018Updated 8 years ago
- Code and data for Learning Rewards from Linguistic Feedback, AAAI '21☆12Dec 16, 2020Updated 5 years ago
- Muesli RL algorithm implementation (PyTorch) (LunarLander-v2)☆20Mar 18, 2024Updated 2 years ago
- deep reinforcement learning using demonstrations to help solve Doom environments☆10Oct 16, 2017Updated 8 years ago
- ☆10Jun 28, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆22Aug 7, 2023Updated 3 years ago
- OpenAI gym environments for goal-conditioned and language-conditioned reinforcement learning☆14Jan 27, 2026Updated 7 months ago
- Catch game example is translated by TensorFlow☆16May 8, 2017Updated 9 years ago
- Multi-Vehicle Movement Sequence Planning algorithm on topological map.☆11Jul 13, 2018Updated 8 years ago
- From pixels to symbolic rule learning☆13Nov 12, 2021Updated 4 years ago
- Poker Simulator☆21Oct 6, 2022Updated 3 years ago
- A SQLModel-based repository dedicated to helping users access the PostgreSQL-based qm9star database more easily in a Python environment.☆14Jun 27, 2025Updated last year
- Unified Model-Free Hierarchical Reinforcement Learning Framework☆37Mar 8, 2019Updated 7 years ago
- This repository is the implementation of the paper "Beating Atari with Natural Language Guided Reinforcement Learning"☆12Nov 25, 2018Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Count based exploration with the successor representation for Unity ML's Pyramid☆13Jun 19, 2019Updated 7 years ago
- 💸 A python library to easily extract information and metrics from Moroccan bank statements.☆13Mar 8, 2023Updated 3 years ago
- PyTorch implementation of Advantage Actor-Critic (A2C)☆47Nov 25, 2017Updated 8 years ago
- ☆14Apr 26, 2018Updated 8 years ago
- Deep Reinforcement Learning by using Truly Proximal Policy Optimization in Tensorflow 2 and Pytorch☆23Nov 9, 2025Updated 10 months ago
- Useful information for new and current MPhil and PhD students in CS @ HKU☆18Jun 14, 2019Updated 7 years ago
- python, ccxt, backtrader, dash☆10Apr 20, 2018Updated 8 years ago
- ☆15Jun 25, 2021Updated 5 years ago
- Environments from the papers "Using Reward Machines for High-Level Task Specification and Decomposition in Reinforcement Learning" and "I…☆11Aug 15, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Deep Q-Network (DQN) with Prioritized Experience Replay (PER)☆17Jan 1, 2020Updated 6 years ago
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- Sentiment analysis system using NLP and machine learning techniques to determine the polarity of the feedback obtained from students in o…☆21Jan 1, 2025Updated last year
- CPSC 330: Applied Machine Learning☆11May 14, 2023Updated 3 years ago
- Implementation of HIRO (Data-Efficient Hierarchical Reinforcement Learning)☆119May 22, 2021Updated 5 years ago
- This is the official repository for the paper "Guided Exploration with Proximal Policy Optimization using a Single Demonstration", https:…☆19Oct 5, 2021Updated 4 years ago
- A tutorial on doing RL research in Julia using both Jupyter notebooks and normal project structures.☆12Jun 23, 2021Updated 5 years ago