Pytorch code for "Learning Belief Representations for Imitation Learning in POMDPs" (UAI 2019)
☆22Aug 4, 2022Updated 4 years ago
Alternatives and similar repositories for BMIL
Users that are interested in BMIL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for the ICML 2020 publication "Information Particle Filter Tree: An Online Algorithm for POMDPs with Belief-Based Rewards on Continu…☆14Jul 3, 2020Updated 6 years ago
- Learning from Guided Play: A Scheduled Hierarchical Approach for Improving Exploration in Adversarial Imitation Learning Source Code☆17Aug 23, 2024Updated last year
- Implicit Normalizing Flows + Reinforcement Learning☆62May 31, 2019Updated 7 years ago
- Deep Variational Reinforcement Learning☆141Jun 21, 2022Updated 4 years ago
- ☆12Dec 22, 2021Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆25Dec 10, 2021Updated 4 years ago
- Domain-Robust Visual Imitation Learning with Mutual Information Constraints code☆19Mar 1, 2021Updated 5 years ago
- Reinforcement learning library for PyTorch.☆11Jun 15, 2018Updated 8 years ago
- Cost-aware Bayesian optimization via the Pandora's box Gittins index☆13Aug 8, 2025Updated last year
- Stochastic Latent Actor-Critic: Deep Reinforcement Learning with a Latent Variable Model☆155Oct 26, 2020Updated 5 years ago
- Official PyTorch Implementation for Metric Residual Networks for Sample Efficient Goal-Conditioned Reinforcement Learning☆21Jan 11, 2023Updated 3 years ago
- Soft Actor-Critic☆162Mar 13, 2018Updated 8 years ago
- Code for the paper "Learning to Do or Learning While Doing: Reinforcement Learning and Bayesian Optimisation for Online Continuous Tuning…☆14Nov 15, 2023Updated 2 years ago
- Pathfinding Using Reinforcement Learning☆12May 21, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Release of NeurIPS 2020 Spotlight paper "Generative Neurosymbolic Machines"☆37Mar 9, 2024Updated 2 years ago
- Implementation of "Active Exploration for Inverse Reinforcement Learning (AceIRL), NeurIPS 2022.☆14Oct 12, 2022Updated 3 years ago
- PyTorch implementation of Stochastic Latent Actor-Critic(SLAC).☆94Jul 25, 2024Updated 2 years ago
- JAX/Haiku implementation of "Auction Learning as a Two-Player Game"☆11Jul 6, 2024Updated 2 years ago
- ☆25Aug 1, 2022Updated 4 years ago
- Multi-agent active perception with prediction rewards☆12Aug 6, 2026Updated last week
- Environments to support https://github.com/sholtodouglas/learning_from_play and reinforcement learning for robotic manipulation.☆21Mar 28, 2021Updated 5 years ago
- 论文Reinforcement Learning of Sequential Price Mechanisms的复现☆12Nov 3, 2022Updated 3 years ago
- Codes for the study "Variational Recurrent Models for Solving Partially Observable Control Tasks", published as a conference paper at ICL…☆55Dec 27, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Implementation of A Distributional Perspective on Reinforcement Learning☆35Aug 1, 2017Updated 9 years ago
- ☆86Apr 10, 2021Updated 5 years ago
- ☆13Mar 12, 2024Updated 2 years ago
- Meta-Inverse Reinforcement Learning with Probabilistic Context Variables☆77Mar 16, 2023Updated 3 years ago
- The PO-UCT algorithm (aka POMCP) implemented in Julia☆39Nov 16, 2025Updated 9 months ago
- Auryn-based simulation of multiplexing and burst-dependent plasticity☆25Feb 18, 2021Updated 5 years ago
- Python package for Dec-POMDP files in the .dpomdp format☆11Oct 28, 2022Updated 3 years ago
- Deep universal probabilistic programming with Python and PyTorch☆14Apr 1, 2020Updated 6 years ago
- code for BINOCULARS and Multi-Step BO☆12Dec 7, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for replicating experiments from the paper, Preference Exploration for Efficient Bayesian Optimization with Multiple Outcomes, publi…☆14Jun 22, 2023Updated 3 years ago
- ☆12Mar 17, 2024Updated 2 years ago
- on-policy optimization baselines for deep reinforcement learning☆32Apr 3, 2020Updated 6 years ago
- NeurIPS 2019 Paper☆12Dec 9, 2019Updated 6 years ago
- Solving POMDPs using exact and approximate methods☆14Aug 9, 2017Updated 9 years ago
- ☆13Apr 11, 2022Updated 4 years ago
- Advantage weighted Actor Critic for Offline RL☆53Aug 27, 2022Updated 3 years ago