Author's PyTorch implementation of SR-DICE for marginalized importance sampling
☆28Dec 7, 2021Updated 4 years ago
Alternatives and similar repositories for SR-DICE
Users that are interested in SR-DICE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for GFlowNet-DPO (Direct Preference Optimization) EMNLP 2024 Main☆20Feb 22, 2026Updated 7 months ago
- Pessimistic Bootstrapping for Uncertainty-Driven Offline Reinforcement Learning☆29Feb 21, 2022Updated 4 years ago
- ☆11Dec 23, 2024Updated last year
- Code associated with our paper "Estimating Risk and Uncertainty in Reinforcement Learning"☆11Oct 3, 2023Updated 2 years ago
- A lightweight reimplementation of Adversarially Trained Actor Critic☆18Mar 19, 2026Updated 6 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆63Jun 6, 2023Updated 3 years ago
- Made for a reading group at the Center for Safe AGI.☆13Feb 23, 2026Updated 7 months ago
- The official implementation of "Mind the Gap: Offline Policy Optimization for Imperfect Rewards" (ICLR2023)☆15Mar 3, 2023Updated 3 years ago
- ☆10Aug 17, 2022Updated 4 years ago
- [ICML 2022] The official implementation of DWBC in "Discriminator-Weighted Offline Imitation Learning from Suboptimal Demonstrations"☆37Jan 5, 2023Updated 3 years ago
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- Author's PyTorch implementation of LAP and PAL with TD3 and DDQN☆44Dec 7, 2021Updated 4 years ago
- ☆11Dec 28, 2023Updated 2 years ago
- ☆17May 25, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- JAX implementation of Graph Attention Networks☆13Jan 29, 2022Updated 4 years ago
- Code for our NeurIPS 2020 paper Improving Generalization in Reinforcement Learning with Mixture Regularization☆34Oct 22, 2020Updated 5 years ago
- Code for the paper "Stabilizing Deep Q-Learning with ConvNets and Vision Transformers under Data Augmentation"☆19Jul 11, 2023Updated 3 years ago
- ☆41Nov 23, 2021Updated 4 years ago
- TensorFlow implementation for our paper "Learning Long-Term Reward Redistribution via Randomized Return Decomposition"☆19Mar 17, 2022Updated 4 years ago
- The official repository of Decoupled Reinforcement Learning to Stabilise Intrinsically-Motivated Exploration" (AAMAS 2022)☆26Feb 3, 2022Updated 4 years ago
- ☆39Jan 1, 2025Updated last year
- ☆29Jan 8, 2026Updated 8 months ago
- Dynamic Simulation Environments for Reinforcement Learning☆13Apr 17, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Code for the paper Novelty Search in Representational Space for Sample Efficient Exploration presented at NeurIPS 2020.☆14Jul 16, 2024Updated 2 years ago
- Learning Laplacian Representations in Reinforcement Learning☆18Jan 2, 2021Updated 5 years ago
- A JAX Implementation of the Twin Delayed DDPG Algorithm☆35Mar 12, 2020Updated 6 years ago
- ☆15Apr 5, 2023Updated 3 years ago
- Pytorch Implementation of AAMAS 2021 paper <Energy-Based Imitation Learning>☆12Oct 8, 2021Updated 4 years ago
- [Preprint] Graph State Space Convolution (GSSC)☆14Jun 11, 2024Updated 2 years ago
- Author's PyTorch implementation of TD7 for online and offline RL☆170Sep 12, 2023Updated 3 years ago
- Challenges and Opportunities in Offline Reinforcement Learning from Visual Observations☆116Apr 16, 2026Updated 5 months ago
- Reinforcement learning algorithms☆41Feb 27, 2019Updated 7 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆19Apr 22, 2024Updated 2 years ago
- Official repository for paper "Versatile Offline Imitation from Observations and Examples via Regularized State-Occupancy Matching" (ICML…☆30Jan 12, 2023Updated 3 years ago
- Library to compare and evaluate reward functions☆69Oct 23, 2023Updated 2 years ago
- Official implementation for "How Should We Meta-Learn Reinforcement Learning Algorithms?"☆23Sep 7, 2025Updated last year
- Decision Mamba: Reinforcement Learning via Sequence Modeling with Selective State Spaces☆54Apr 1, 2024Updated 2 years ago
- Train, evaluate, and optimize implicit feedback-based recommender systems.☆31Sep 8, 2026Updated 2 weeks ago
- Representation Learning in RL☆13Jun 1, 2022Updated 4 years ago