Simple bit flipping with sparse rewards using HER, similarly to the original paper
☆39Feb 25, 2019Updated 7 years ago
Alternatives and similar repositories for Hindsight-Experience-Replay---Bit-Flipping
Users that are interested in Hindsight-Experience-Replay---Bit-Flipping are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Hindsight Experience Replay - Bit flipping experiment in Tensorflow☆58Oct 7, 2018Updated 7 years ago
- ☆13Apr 4, 2023Updated 3 years ago
- ☆16Oct 3, 2023Updated 2 years ago
- rlcourse-march-17-hugobb created by GitHub Classroom☆15Jul 3, 2024Updated 2 years ago
- Logarithmic Reinforcement Learning☆28Apr 7, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Comp 781 Project☆10Jan 2, 2026Updated 7 months ago
- ☆13Feb 3, 2021Updated 5 years ago
- Implementation based on the paper "Ant Colony System for Graph Coloring Problem"☆11Apr 14, 2018Updated 8 years ago
- deep reinforcement learning using demonstrations to help solve Doom environments☆10Oct 16, 2017Updated 8 years ago
- ☆10Jun 28, 2022Updated 4 years ago
- ☆22Aug 7, 2023Updated 3 years ago
- OpenAI gym environments for goal-conditioned and language-conditioned reinforcement learning☆14Jan 27, 2026Updated 6 months ago
- Catch game example is translated by TensorFlow☆16May 8, 2017Updated 9 years ago
- From pixels to symbolic rule learning☆12Nov 12, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Poker Simulator☆20Oct 6, 2022Updated 3 years ago
- ☆11Jul 29, 2021Updated 5 years ago
- This repository is the implementation of the paper "Beating Atari with Natural Language Guided Reinforcement Learning"☆12Nov 25, 2018Updated 7 years ago
- Count based exploration with the successor representation for Unity ML's Pyramid☆12Jun 19, 2019Updated 7 years ago
- Deep Reinforcement Learning by using Truly Proximal Policy Optimization in Tensorflow 2 and Pytorch☆22Nov 9, 2025Updated 9 months ago
- Python Implementation of a simplified version of STICS crop model☆16Sep 12, 2024Updated last year
- python, ccxt, backtrader, dash☆10Apr 20, 2018Updated 8 years ago
- Environments from the papers "Using Reward Machines for High-Level Task Specification and Decomposition in Reinforcement Learning" and "I…☆12Aug 15, 2023Updated 2 years ago
- Official codes for "Training Deep Q-Network via Monte Carlo Tree Search for Adaptive Bitrate Control in Video Delivery"☆10Jul 21, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Training code for the ACAM action detection model.☆29Feb 2, 2023Updated 3 years ago
- This is the official repository for the paper "Guided Exploration with Proximal Policy Optimization using a Single Demonstration", https:…☆19Oct 5, 2021Updated 4 years ago
- Finetuning VITS Efficiently☆32Nov 6, 2023Updated 2 years ago
- Reimplementation of Policy Optimization with Demonstrations (POfD) from ICML 2018.☆16Jun 5, 2019Updated 7 years ago
- Official code for Conformal Isometry of Lie Group Representation in Recurrent Network of Grid Cells (NeurIPS workshop on Symmetry and Geo…☆13Nov 1, 2022Updated 3 years ago
- This software calculates the Energy Yield of single and multi-junction solar cells. It consists of individual modules taking care of deri…☆19Aug 31, 2022Updated 3 years ago
- An environment for tabular Reinforcement Learning agents.☆14Jun 13, 2018Updated 8 years ago
- DRL for WebRTC Control☆13Feb 3, 2024Updated 2 years ago
- 使用numpy从零开始实现llama3的推理流程,并对其进行封装,对比GPU,CPU上的表现以及Lora微调。llama3 implemented from scratch using numpy and lora fine-tune.。☆11Jul 16, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- VITS-based zero-shot TTS system varying with diverse style/speaker conditioning methods.☆36Sep 21, 2022Updated 3 years ago
- ☆10Apr 24, 2021Updated 5 years ago
- Distributional Successor Features Enable Zero-Shot Policy Optimization☆15Apr 11, 2025Updated last year
- Machine Learning approaches to predict Direct Horizontal and Direct Normal Irradiance from the sun.☆14Nov 20, 2020Updated 5 years ago
- Deep reinforcement learning baselines base on OpenAI. More algorithms are included, such as Rainbow: Combining Improvements in Deep Rei…☆35Aug 23, 2018Updated 7 years ago
- The framework to deal with ctr problem。The project contains FNN,PNN,DEEPFM, NFM etc☆18Feb 6, 2018Updated 8 years ago
- Your one stop CLI for ONNX model analysis.☆48Nov 13, 2022Updated 3 years ago