Code for Posterior Sampling for Deep Reinforcement Learning, ICML 2023
☆28Mar 7, 2024Updated 2 years ago
Alternatives and similar repositories for PSDRL
Users that are interested in PSDRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Jul 4, 2022Updated 4 years ago
- Code for the NeurIPS 2021 paper "Deep Bandits Show-Off: Simple and Efficient Exploration with Deep Networkst"☆14Sep 12, 2022Updated 3 years ago
- Official implementation of Harnessing Mixed Offline Reinforcement Learning Datasets via Trajectory Reweighting☆16Feb 14, 2024Updated 2 years ago
- Minimal Decision Transformer Implementation written in Jax (Flax).☆18Aug 8, 2022Updated 3 years ago
- iQRL: implicitly Quantized Representations for Sample-efficient Reinforcement Learning☆12Jan 8, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Accompanying Code for "Flipping Coins to Estimate Pseudocounts for Exploration in Reinforcement Learning", ICML 2023☆25Dec 29, 2023Updated 2 years ago
- Official Implementation of NeurIPS'23 Paper "Cross-Episodic Curriculum for Transformer Agents"☆32Oct 12, 2023Updated 2 years ago
- IV-RL - Sample Efficient Deep Reinforcement Learning via Uncertainty Estimation☆40Jul 18, 2025Updated last year
- Code release for Efficient Planning in a Compact Latent Action Space (ICLR2023) https://arxiv.org/abs/2208.10291.☆113May 12, 2023Updated 3 years ago
- Integrate AutoRL into DQN to implement a single traffic signal control system.☆16Nov 16, 2023Updated 2 years ago
- This repository is the official implementation of Bidirectional Learning for Offline Infinite-width Model-based Optimization (NeurIPS 202…☆14Jan 19, 2023Updated 3 years ago
- Code for☆15Oct 16, 2020Updated 5 years ago
- Code Release for Task Agnostic Dynamics Priors for Deep Reinforcement Learning☆12Jun 13, 2019Updated 7 years ago
- [ICML2026] Official JAX code for Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying☆15Jul 3, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆18Apr 17, 2026Updated 3 months ago
- ☆26Jan 26, 2024Updated 2 years ago
- code for "Decoupled Preference-based Reinforcement Learning for Personalized Human-Robot Interaction"☆11Jul 9, 2022Updated 4 years ago
- ☆12Apr 25, 2022Updated 4 years ago
- Code implementation of "Information Design in Multi-Agent Reinforcement Learning"☆16Aug 18, 2023Updated 2 years ago
- Model Predictive Control-based Reinforcement Learning with Control Barrier Functions☆29Jan 16, 2026Updated 6 months ago
- Learning from Guided Play: A Scheduled Hierarchical Approach for Improving Exploration in Adversarial Imitation Learning Source Code☆17Aug 23, 2024Updated last year
- Author's Pytorch implementation of ICLR2023 paper Behavior Proximal Policy Optimization (BPPO).☆94Dec 13, 2023Updated 2 years ago
- [ICLR 2024] Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.☆25Apr 19, 2024Updated 2 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Generative cellular automaton-like learning environments for RL.☆20Jan 30, 2025Updated last year
- Deep reinforcement approach to solving dynamic pickup and delivery problem☆21Jun 7, 2021Updated 5 years ago
- Author's implementation of ReBRAC, a minimalist improvement upon TD3+BC☆19Oct 22, 2023Updated 2 years ago
- This repository contains PyTorch implementations of deep reinforcement learning algorithms and environments for Robotics and Controls. T…☆19Mar 20, 2022Updated 4 years ago
- A Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime☆15Dec 7, 2024Updated last year
- using monte carlo dropout to have uncertainty estimation of predictions☆16Nov 12, 2019Updated 6 years ago
- ☆18Jul 10, 2022Updated 4 years ago
- A project to assess the costs of flexible vehicle routing strategies☆20Apr 24, 2023Updated 3 years ago
- Simulation of Ridesharing Market and the MDP Order Dispatch Policy☆21Mar 13, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- MEAformer: An all-MLP Transformer with Temporal External Attention for Long-term Time Series Forecasting☆13Apr 27, 2024Updated 2 years ago
- ☆11May 29, 2023Updated 3 years ago
- Human - Robot Collaboration for fabric folding using Kinect2, RoboDK, Reflex 1 gripper and the ATI Force Torque Gamma sensor☆15Mar 1, 2023Updated 3 years ago
- Pessimistic Value Iteration for Multi-Task Data Sharing in Offline RL☆18Nov 21, 2023Updated 2 years ago
- Official code for ICML 2024 paper Reinformer: Max-Return Sequence Modeling for offline RL☆49Oct 16, 2024Updated last year
- Analyzing different stocks listed on the NASDAQ stock market☆13Dec 5, 2020Updated 5 years ago
- ☆12May 15, 2026Updated 2 months ago