Code for Posterior Sampling for Deep Reinforcement Learning, ICML 2023
☆28Mar 7, 2024Updated 2 years ago
Alternatives and similar repositories for PSDRL
Users that are interested in PSDRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Jul 4, 2022Updated 4 years ago
- Code for the NeurIPS 2021 paper "Deep Bandits Show-Off: Simple and Efficient Exploration with Deep Networkst"☆14Sep 12, 2022Updated 3 years ago
- This is pytorch implmentation project of Bootsrapped DQN☆13Dec 6, 2020Updated 5 years ago
- Official implementation of Harnessing Mixed Offline Reinforcement Learning Datasets via Trajectory Reweighting☆16Feb 14, 2024Updated 2 years ago
- Minimal Decision Transformer Implementation written in Jax (Flax).☆18Aug 8, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Efficient Exploration through Bayesian Deep Q-Networks☆38Feb 14, 2018Updated 8 years ago
- iQRL: implicitly Quantized Representations for Sample-efficient Reinforcement Learning☆12Jan 8, 2025Updated last year
- Single-file SAC-N implementation on jax with flax and equinox. 10x faster than pytorch☆57May 21, 2023Updated 3 years ago
- Accompanying Code for "Flipping Coins to Estimate Pseudocounts for Exploration in Reinforcement Learning", ICML 2023☆25Dec 29, 2023Updated 2 years ago
- [ICML 2023] Official code for "DevFormer: A Symmetric Transformer for Context-Aware Device Placement"☆23Dec 7, 2024Updated last year
- Official Implementation of NeurIPS'23 Paper "Cross-Episodic Curriculum for Transformer Agents"☆32Oct 12, 2023Updated 2 years ago
- Code release for Efficient Planning in a Compact Latent Action Space (ICLR2023) https://arxiv.org/abs/2208.10291.☆113May 12, 2023Updated 3 years ago
- Integrate AutoRL into DQN to implement a single traffic signal control system.☆16Nov 16, 2023Updated 2 years ago
- Recorder for Azure Kinect☆14Aug 6, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This repository is the official implementation of Bidirectional Learning for Offline Infinite-width Model-based Optimization (NeurIPS 202…☆14Jan 19, 2023Updated 3 years ago
- Utilities for visualizing the human poses of CHICO dataset from "Pose Forecasting in Industrial Human-Robot Collaboration" ECCV 2022 pape…☆11Oct 24, 2022Updated 3 years ago
- CabiNet: Scaling Object Rearrangement in Clutter☆24Jan 17, 2024Updated 2 years ago
- Code Release for Task Agnostic Dynamics Priors for Deep Reinforcement Learning☆12Jun 13, 2019Updated 7 years ago
- Code for☆16Oct 16, 2020Updated 5 years ago
- Implementation of NeurIPS 2018 paper "Meta-Gradient Reinforcement Learning"☆21Jul 19, 2022Updated 4 years ago
- [ICML2026] Official JAX code for Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying☆17Jul 3, 2026Updated last month
- Mamo: a Mathematical Modeling Benchmark with Solvers☆15Jun 12, 2024Updated 2 years ago
- ☆14Apr 1, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆18Apr 17, 2026Updated 4 months ago
- ☆22Sep 22, 2022Updated 3 years ago
- ☆26Jan 26, 2024Updated 2 years ago
- code for "Decoupled Preference-based Reinforcement Learning for Personalized Human-Robot Interaction"☆11Jul 9, 2022Updated 4 years ago
- ☆12Apr 25, 2022Updated 4 years ago
- Online Resource Repository: Datasets, Simulation Platforms, and Empirical Research on Emerging Mixed Traffic of Automated Vehicles and Hu…☆18Nov 29, 2023Updated 2 years ago
- Code implementation of "Information Design in Multi-Agent Reinforcement Learning"☆16Aug 18, 2023Updated 3 years ago
- Model Predictive Control-based Reinforcement Learning with Control Barrier Functions☆32Jan 16, 2026Updated 7 months ago
- Learning from Guided Play: A Scheduled Hierarchical Approach for Improving Exploration in Adversarial Imitation Learning Source Code☆17Aug 23, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Author's Pytorch implementation of ICLR2023 paper Behavior Proximal Policy Optimization (BPPO).☆96Dec 13, 2023Updated 2 years ago
- [ICLR 2024] Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View.☆25Apr 19, 2024Updated 2 years ago
- Author's implementation of ReBRAC, a minimalist improvement upon TD3+BC☆19Oct 22, 2023Updated 2 years ago
- This repository contains PyTorch implementations of deep reinforcement learning algorithms and environments for Robotics and Controls. T…☆19Mar 20, 2022Updated 4 years ago
- A Benchmark Platform for Reinforcement Learning Based Dynamic Treatment Regime☆14Dec 7, 2024Updated last year
- using monte carlo dropout to have uncertainty estimation of predictions☆16Nov 12, 2019Updated 6 years ago
- Simulation of Ridesharing Market and the MDP Order Dispatch Policy☆21Mar 13, 2024Updated 2 years ago