G-HER algorithm
☆18May 24, 2019Updated 7 years ago
Alternatives and similar repositories for GHER
Users that are interested in GHER are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Modular-HER is revised from OpenAI baselines and supports many improvements for Hindsight Experience Replay as modules.☆17Jun 23, 2021Updated 5 years ago
- Maximum Entropy-Regularized Multi-Goal Reinforcement Learning (ICML 2019)☆24May 30, 2019Updated 7 years ago
- Energy-Based Hindsight Experience Prioritization (CoRL 2018) Oral presentation (7%)☆35Nov 28, 2018Updated 7 years ago
- Reinforcement learning algorithms implemented to learn to play Super Mario Bros.☆12Oct 24, 2020Updated 5 years ago
- Author implementation of Monte Carlo Augmented Actor Critic in PyTorch☆18Oct 24, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- DHER: Hindsight Experience Replay for Dynamic Goals (ICLR-2019)☆65Nov 8, 2019Updated 6 years ago
- Simulation source code and examples for applying machine learning on Sawyer Robot☆11Dec 11, 2018Updated 7 years ago
- A2C training of Relational Deep Reinforcement Learning Architecture☆13Jun 22, 2022Updated 4 years ago
- Codebase for "Causal Induction from Visual Observations for Goal-Directed Tasks"☆14Feb 25, 2020Updated 6 years ago
- Blog post: how to do deterministic policy gradient with gumbel softmax and why you should do it.☆12Jun 20, 2017Updated 9 years ago
- Object Detection (Tensorflow) and horizon line detection (OpenCV) for drone images☆11Feb 6, 2018Updated 8 years ago
- ☆18Oct 7, 2019Updated 6 years ago
- Python Q learning implementations and application examples in wireless networks☆19Feb 10, 2017Updated 9 years ago
- Implementation of DDPG+HER on gym robotics environment FetchReach-v1☆33Nov 13, 2018Updated 7 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Hindsight policy gradients☆46Jan 31, 2020Updated 6 years ago
- Tensorflow eager implementation of Pix2Pix (Image-to-image translation with conditional adversarial networks)☆12Aug 12, 2019Updated 7 years ago
- Trajectory planning based on RL with Hindsight Experience Replay & Dense Reward Engineering to solve openai-gym robotics "FetchReach-v1" …☆19May 12, 2022Updated 4 years ago
- ☆11Oct 26, 2022Updated 3 years ago
- ☆18Jan 3, 2022Updated 4 years ago
- Python bindings to some optimization benchmarks (robotics problems), in order to constrained optimization solvers. Includes also an inter…☆21Dec 8, 2022Updated 3 years ago
- Code for Sufficient Input Subsets Paper☆14Mar 8, 2019Updated 7 years ago
- Intelligent control algorithm and simulation environment.☆17Jan 6, 2020Updated 6 years ago
- ☆22Mar 4, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is the pytorch implementation of Hindsight Experience Replay (HER) - Experiment on all fetch robotic environments.☆451Dec 11, 2021Updated 4 years ago
- Algorithms described in the paper Hindsight Credit Assignment (NeurIPS 2019).☆11Oct 27, 2019Updated 6 years ago
- MuJoCo benchmark for Deep Reinforcement Learning as provided by Tianshou framework.☆14Jan 12, 2025Updated last year
- ☆12Dec 8, 2022Updated 3 years ago
- Pessimistic Value Iteration for Multi-Task Data Sharing in Offline RL☆18Nov 21, 2023Updated 2 years ago
- Deep Latent Gamma Model / Gamma VAE☆14May 9, 2018Updated 8 years ago
- Modified Fetch Robotics environments from OpenAI gym☆11Nov 27, 2021Updated 4 years ago
- Code repository of NeurIPS 2021 paper: Detecting and Adapting to Irregular Distribution Shifts in Bayesian Online Learning.☆12Oct 17, 2022Updated 3 years ago
- A pytorch-based implementation of Dirichlet Process Mixture Model (DPMM)☆13Apr 20, 2026Updated 3 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- LEGO : Leveraging Experience with Graph Oracles☆11Feb 15, 2020Updated 6 years ago
- Online Preference Alignment for Language Models via Count-based Exploration☆21Jan 14, 2025Updated last year
- ppo+action mask for atari tennis agent☆12Mar 2, 2023Updated 3 years ago
- Implementation of Relational Deep Reinforcement Learning☆26Jan 31, 2020Updated 6 years ago
- ☆16Sep 25, 2019Updated 6 years ago
- Code for our NeurIPS 2020 paper Improving Generalization in Reinforcement Learning with Mixture Regularization☆34Oct 22, 2020Updated 5 years ago
- ☆14Oct 27, 2019Updated 6 years ago