Author's PyTorch implementation of SR-DICE for marginalized importance sampling
☆28Dec 7, 2021Updated 4 years ago
Alternatives and similar repositories for SR-DICE
Users that are interested in SR-DICE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for GFlowNet-DPO (Direct Preference Optimization) EMNLP 2024 Main☆20Feb 22, 2026Updated 5 months ago
- ☆19Jan 2, 2024Updated 2 years ago
- ☆15Dec 10, 2025Updated 8 months ago
- Pessimistic Bootstrapping for Uncertainty-Driven Offline Reinforcement Learning☆29Feb 21, 2022Updated 4 years ago
- A python package for loading robotics datasets which were recorded on the TriFinger platform. Also contains simulated gym environments th…☆17Jan 17, 2024Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆11Dec 23, 2024Updated last year
- Code associated with our paper "Estimating Risk and Uncertainty in Reinforcement Learning"☆11Oct 3, 2023Updated 2 years ago
- ☆61Jun 6, 2023Updated 3 years ago
- A lightweight reimplementation of Adversarially Trained Actor Critic☆19Mar 19, 2026Updated 4 months ago
- Made for a reading group at the Center for Safe AGI.☆12Feb 23, 2026Updated 5 months ago
- ☆17Dec 31, 2024Updated last year
- The official implementation of "Mind the Gap: Offline Policy Optimization for Imperfect Rewards" (ICLR2023)☆15Mar 3, 2023Updated 3 years ago
- ☆10Aug 17, 2022Updated 4 years ago
- [ICML 2022] The official implementation of DWBC in "Discriminator-Weighted Offline Imitation Learning from Suboptimal Demonstrations"☆37Jan 5, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- Author's PyTorch implementation of LAP and PAL with TD3 and DDQN☆43Dec 7, 2021Updated 4 years ago
- ☆11Dec 28, 2023Updated 2 years ago
- ☆17May 25, 2023Updated 3 years ago
- JAX implementation of Graph Attention Networks☆13Jan 29, 2022Updated 4 years ago
- Code for our NeurIPS 2020 paper Improving Generalization in Reinforcement Learning with Mixture Regularization☆34Oct 22, 2020Updated 5 years ago
- Code for the paper "Stabilizing Deep Q-Learning with ConvNets and Vision Transformers under Data Augmentation"☆19Jul 11, 2023Updated 3 years ago
- ☆40Nov 23, 2021Updated 4 years ago
- TensorFlow implementation for our paper "Learning Long-Term Reward Redistribution via Randomized Return Decomposition"☆19Mar 17, 2022Updated 4 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- The official repository of Decoupled Reinforcement Learning to Stabilise Intrinsically-Motivated Exploration" (AAMAS 2022)☆26Feb 3, 2022Updated 4 years ago
- Tutorials on learning and using successor representations.☆55Oct 31, 2019Updated 6 years ago
- ☆39Jan 1, 2025Updated last year
- ☆16Oct 5, 2021Updated 4 years ago
- ☆29Jan 8, 2026Updated 7 months ago
- Dynamic Simulation Environments for Reinforcement Learning☆13Apr 17, 2021Updated 5 years ago
- The official code release for "Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning", ICLR 2025☆22May 28, 2025Updated last year
- Code for the paper Novelty Search in Representational Space for Sample Efficient Exploration presented at NeurIPS 2020.☆14Jul 16, 2024Updated 2 years ago
- Official code for "Pretraining Representations For Data-Efficient Reinforcement Learning" (NeurIPS 2021)☆56Jul 27, 2021Updated 5 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A JAX Implementation of the Twin Delayed DDPG Algorithm☆35Mar 12, 2020Updated 6 years ago
- Pytorch Implementation of AAMAS 2021 paper <Energy-Based Imitation Learning>☆12Oct 8, 2021Updated 4 years ago
- [Preprint] Graph State Space Convolution (GSSC)☆14Jun 11, 2024Updated 2 years ago
- Author's PyTorch implementation of TD7 for online and offline RL☆170Sep 12, 2023Updated 2 years ago
- Challenges and Opportunities in Offline Reinforcement Learning from Visual Observations☆115Apr 16, 2026Updated 4 months ago
- Reinforcement learning algorithms☆41Feb 27, 2019Updated 7 years ago
- ☆19Apr 22, 2024Updated 2 years ago