Code repo for Gradient Temporal-Difference Learning with Regularized Corrections paper.
☆38Oct 14, 2020Updated 5 years ago
Alternatives and similar repositories for Regularized-GradientTD
Users that are interested in Regularized-GradientTD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Nov 9, 2021Updated 4 years ago
- Experiment utility code, specifically designed for use with Compute Canada.☆11Jan 27, 2025Updated last year
- This repository contains the code used in the paper Evaluating the Performance of Reinformcent Learning Algorithms☆27Aug 14, 2021Updated 4 years ago
- (ICLR 2021) Learning to Represent Action Values as a Hypergraph on the Action Vertices☆23Jun 22, 2021Updated 5 years ago
- ☆27Mar 11, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Easy MDPs and grid worlds with accessible transition dynamics to do exact calculations☆49Apr 1, 2022Updated 4 years ago
- Reinforcement learning algorithms☆41Feb 27, 2019Updated 7 years ago
- Network Randomization: A Simple Technique for Generalization in Deep Reinforcement Learning / ICLR 2020☆57Apr 27, 2020Updated 6 years ago
- This is a project using Pytorch to fulfill reinforcement learning on a simple game - Gridworld☆14Jul 13, 2020Updated 6 years ago
- Python library for finding English tenses in sentences☆12Jun 9, 2019Updated 7 years ago
- ☆10Apr 24, 2021Updated 5 years ago
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- A graduate-level introduction to reinforcement learning as a framework for modeling, optimization, and control, connecting dynamic models…☆18Dec 9, 2025Updated 7 months ago
- ☆333Dec 19, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Authors' PyTorch implementation of 'Recomposing the Reinforcement Learning Building-Blocks with Hypernetworks' (HypeRL)☆26Jun 9, 2021Updated 5 years ago
- ☆14May 30, 2019Updated 7 years ago
- Unified notation for Markov Decision Processes PO(MDP)s☆24Apr 27, 2018Updated 8 years ago
- Implementation prototype of the Deep Deterministic Off-Policy Gradient (DD-OPG) method.☆11Jun 12, 2019Updated 7 years ago
- ☆29Jun 27, 2026Updated 3 weeks ago
- Implementation of NeurIPS2021 paper <On Effective Scheduling of Model-based Reinforcement Learning>☆13Nov 16, 2021Updated 4 years ago
- TensorFlow implementation of "noisy K-FAC" and "noisy EK-FAC".☆61Jan 12, 2019Updated 7 years ago
- ☆16Jun 1, 2023Updated 3 years ago
- Official implementation for the paper: "Shallow Updates for Deep Reinforcement Learning"☆18Nov 2, 2017Updated 8 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Revisiting Peng's Q(lambda) for Modern Reinforcement Learning☆15Jul 23, 2021Updated 5 years ago
- Computes trajectories for evolutionary dynamics.☆15Oct 6, 2020Updated 5 years ago
- ☆33Mar 19, 2024Updated 2 years ago
- Retrieve information from DBLP and update BibTex files automatically☆53Jun 4, 2022Updated 4 years ago
- Automatic Data-Regularized Actor-Critic (Auto-DrAC)☆104Mar 24, 2023Updated 3 years ago
- Official repo for our AAAI'21 paper, https://arxiv.org/abs/2007.12354☆30Jul 14, 2021Updated 5 years ago
- IV-RL - Sample Efficient Deep Reinforcement Learning via Uncertainty Estimation☆40Jul 18, 2025Updated last year
- Implementation of the skill discovery algorithm described in ICLR submission "Option Discovery using Deep Skill Chaining"☆30Sep 24, 2019Updated 6 years ago
- Mirror Descent Policy Optimization☆43Oct 31, 2020Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code to reproduce the experiments in The Mirage of Action-Dependent Baselines in Reinforcement Learning.☆17Aug 2, 2018Updated 7 years ago
- some common TD Learning algorithms☆66Mar 6, 2020Updated 6 years ago
- Code for "Best arm identification in multi-armed bandits with delayed feedback", AISTATS 2018.☆20Apr 3, 2018Updated 8 years ago
- krazy grid world☆26Mar 2, 2020Updated 6 years ago
- This repository releases the code and data for utterance rewriting in open-domain dialogues.☆18Feb 24, 2023Updated 3 years ago
- Multi-view Reinforcement Learning☆11Feb 9, 2020Updated 6 years ago
- OpenAI Gym Wrapper for DeepMind Control Suite☆74Nov 30, 2021Updated 4 years ago