[ICLR 22] Value Gradient weighted Model-Based Reinforcement Learning.
☆26Apr 15, 2023Updated 3 years ago
Alternatives and similar repositories for vagram
Users that are interested in vagram are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- JAX code for the paper "Control-Oriented Model-Based Reinforcement Learning with Implicit Differentiation"☆43Jun 14, 2021Updated 5 years ago
- The Official Code for Offline Model-based Adaptable Policy Learning (NeurIPS'21 & TPAMI)☆25Jan 16, 2024Updated 2 years ago
- ☆14Sep 14, 2020Updated 6 years ago
- Learning Action-Value Gradients in Model-based Policy Optimization☆32Sep 7, 2021Updated 5 years ago
- Simplifying Model-based RL: Learning Representations, Latent-space Models and Policies with One Objective☆82Mar 9, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Efficient seed-parallel implementation of "Breaking the Replay Ratio Barrier"☆29May 22, 2023Updated 3 years ago
- Author's PyTorch implementation of Randomized Ensembled Double Q-Learning (REDQ) algorithm.☆188Nov 14, 2024Updated last year
- ☆17May 25, 2023Updated 3 years ago
- ☆11Oct 14, 2019Updated 6 years ago
- A PyTorch implementation of visual interaction networks☆12Jul 1, 2019Updated 7 years ago
- Code for the paper "When to Trust Your Model: Model-Based Policy Optimization"☆562Nov 22, 2022Updated 3 years ago
- Public Release of Plan2vec Implementation in pyTorch☆57Oct 28, 2022Updated 3 years ago
- Plannable Approximations to MDP Homomorphisms: Equivariance under Actions☆30Jun 30, 2020Updated 6 years ago
- Windy GridWorlds environments compatible with OpenAI gym.☆15Jul 8, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆99Mar 24, 2023Updated 3 years ago
- Code for reproducing experiments in Model-Based Active Exploration, ICML 2019☆81Jul 23, 2019Updated 7 years ago
- Code for NeurIPS 2021 paper "Curriculum Offline Imitation Learning"☆18Oct 21, 2022Updated 3 years ago
- [ICLR 2025] Bootstrapped Model Predictive Control☆41Jul 20, 2026Updated 2 months ago
- ☆18Feb 7, 2021Updated 5 years ago
- This repository collects supplementary material to study reinforcement learning with a focus on topics covered by the TU Darmstadt IAS le…☆12Feb 11, 2019Updated 7 years ago
- Neural Fixed-Point Acceleration for Convex Optimization☆30Oct 6, 2022Updated 4 years ago
- Library that provides environments for planning problems☆17Apr 24, 2026Updated 5 months ago
- Estimating Q(s,s') with Deep Deterministic Dynamics Gradients☆32Feb 21, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- On the model-based stochastic value gradient for continuous reinforcement learning☆58Mar 6, 2026Updated 7 months ago
- ☆14Oct 20, 2020Updated 5 years ago
- ☆30Mar 1, 2022Updated 4 years ago
- Simulation system for path planning evaluation☆13Dec 13, 2025Updated 9 months ago
- PyTorch implementation of Stochastic Latent Actor-Critic(SLAC).☆94Jul 25, 2024Updated 2 years ago
- Semi-Supervised Offline Reinforcement Learning with Action-Free Trajectories☆41Jul 16, 2023Updated 3 years ago
- Accompanying code for "Learning and Planning in Average-Reward Markov Decision Processes"☆15Feb 10, 2021Updated 5 years ago
- Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees☆93Sep 13, 2019Updated 7 years ago
- ☆12Mar 14, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Control of a hovering AUV simulated using HoloOcean☆58Sep 28, 2023Updated 3 years ago
- ☆15Jun 8, 2023Updated 3 years ago
- Code for the paper "Gamma-Models: Generative Temporal Difference Learning for Infinite-Horizon Prediction"☆48Sep 20, 2023Updated 3 years ago
- Gantry provides an API that streamlines running experiments in Beaker☆33Jul 20, 2026Updated 2 months ago
- Ant Gather and Ant Maze envs, separated from RLLab☆11Aug 2, 2018Updated 8 years ago
- CIC: Contrastive Intrinsic Control for Unsupervised Skill Discovery☆88Jul 27, 2022Updated 4 years ago
- Library for Model Based RL☆1,063Jul 12, 2024Updated 2 years ago