Implementation of (Learning Continuous Control Policies by Stochastic Value Gradients)[https://arxiv.org/abs/1510.09142]
☆25Jan 15, 2022Updated 4 years ago
Alternatives and similar repositories for stochastic_value_gradient
Users that are interested in stochastic_value_gradient are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆69May 26, 2018Updated 8 years ago
- Learning Backtracking Models, ICLR'19☆10Feb 2, 2018Updated 8 years ago
- Learning Action-Value Gradients in Model-based Policy Optimization☆32Sep 7, 2021Updated 5 years ago
- RobustStabilityGuaranteeRL☆10Aug 22, 2019Updated 7 years ago
- Revisiting Peng's Q(lambda) for Modern Reinforcement Learning☆15Jul 23, 2021Updated 5 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Reinforcement Leanring Algorithms Trained with Unity☆13Apr 26, 2019Updated 7 years ago
- (CoRL 2019 Spotlight) Asynchronous Methods for Model-Based Reinforcement Learning☆14Dec 27, 2022Updated 3 years ago
- 现在好用的能同步的网盘都没有了,于是自己用阿里云的OSS撸了一个☆10Aug 24, 2026Updated 3 weeks ago
- ☆10Apr 18, 2017Updated 9 years ago
- Self-Consistent Trajectory Autoencoder: Hierarchical Reinforcement Learning with Trajectory Embeddings☆96Jun 8, 2018Updated 8 years ago
- Official implementation of DynE, Dynamics-aware Embeddings for RL☆45Apr 28, 2021Updated 5 years ago
- Open AI gym environment for the Baxter robot☆14Oct 6, 2016Updated 9 years ago
- ☆14Oct 20, 2020Updated 5 years ago
- Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees☆93Sep 13, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ICRL 2020☆20Feb 18, 2020Updated 6 years ago
- Intrinsic Motivation and Automatic Curricula via Asymmetric Self-Play