Model-Based Uncertainty in Value Functions (AISTATS2023)
☆16Feb 28, 2023Updated 3 years ago
Alternatives and similar repositories for ube-mbrl
Users that are interested in ube-mbrl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official pytorch implementation for our ICLR 2023 paper "Latent State Marginalization as a Low-cost Approach for Improving Exploration".☆24Feb 9, 2023Updated 3 years ago
- ☆26Jan 26, 2024Updated 2 years ago
- Code for ICLR 2022 Paper (HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning)☆12Nov 28, 2023Updated 2 years ago
- Code accompanying the paper "Information Directed Reward Learning for Reinforcement Learning" (NeurIPS 2021).☆13Nov 16, 2021Updated 4 years ago
- ☆18Oct 15, 2020Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Author's PyTorch implementation of ICML'23 paper "Policy Regularization with Dataset Constraint for Offline Reinforcement Learning" for D…☆18Nov 8, 2024Updated last year
- Implementation of Neurips 2023 Paper "Multi Time Scale World Models"☆18Nov 8, 2024Updated last year
- [ICLR 2024 oral] Pre-Training Goal-based Models for Sample-Efficient Reinforcement Learning☆30Mar 1, 2024Updated 2 years ago
- Clean, extensible implementation of MACAW [ICML 2021]☆12Dec 7, 2021Updated 4 years ago
- ☆36Jul 10, 2026Updated 2 weeks ago
- Implementation of "Active Exploration for Inverse Reinforcement Learning (AceIRL), NeurIPS 2022.☆14Oct 12, 2022Updated 3 years ago
- Code to accompany the paper "Mismatched No More: Joint Model-Policy Optimization for Model-Based RL"☆21Oct 6, 2021Updated 4 years ago
- A minimal home grid world environment to evaluate language understanding in interactive agents.☆24Sep 6, 2023Updated 2 years ago
- Learning Laplacian Representations in Reinforcement Learning☆18Jan 2, 2021Updated 5 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation of Tactical Optimistic and Pessimistic value estimation☆25Jul 18, 2023Updated 3 years ago
- Official Implementation of NeurIPS'23 Paper "Cross-Episodic Curriculum for Transformer Agents"☆32Oct 12, 2023Updated 2 years ago
- A flexible computer-vision pipeline that connects to video source, detects, tracks and geo-maps objects.☆11Jul 9, 2026Updated 2 weeks ago
- (RSS 2021) Move Beyond Trajectories: Distribution Space Coupling for Crowd Navigation☆23Feb 9, 2025Updated last year
- Official code for "Pretraining Representations For Data-Efficient Reinforcement Learning" (NeurIPS 2021)☆56Jul 27, 2021Updated 5 years ago
- video prediction and world model research☆14Jun 10, 2022Updated 4 years ago
- ☆25Feb 21, 2022Updated 4 years ago
- Tracking literature and additional online resources on transformers for sequential decision making including RL and beyond.☆52Dec 21, 2022Updated 3 years ago
- Author's implementation of ReBRAC, a minimalist improvement upon TD3+BC☆63Aug 3, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆11Nov 18, 2023Updated 2 years ago
- Geometry generation/planning for robotically assembled spatial structures☆14Mar 23, 2023Updated 3 years ago
- Exact nearest neighbor searching for various Euclidean, SO(3), SE(3) and weighted combinations thereof.☆11Oct 5, 2018Updated 7 years ago
- Logging library for JAX that is compatible with transformations and primitives such as vmap and scan.☆16Jul 16, 2026Updated last week
- "Control is as much an effect as a cause, and the idea that control is something you exert is a real handicap to progress." ― Steve Grand☆12Apr 30, 2020Updated 6 years ago
- ☆17Apr 23, 2026Updated 3 months ago
- A clean, extensible toolbox for running experiments on the Crazyflie 2.0 quadrotor.☆17Feb 23, 2024Updated 2 years ago
- IRIS plans motions that allow a robot to inspect a set of points of interest (POIs), aiming at maximizing the number of POI inspected wit…☆13Apr 5, 2024Updated 2 years ago
- ☆65Jan 30, 2026Updated 5 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- PyRoboCOP is a lightweight Python-based package for control and optimization of robotic systems described by nonlinear Differential Algeb…☆62Jan 25, 2023Updated 3 years ago
- ☆20Jun 25, 2023Updated 3 years ago
- Planner Developer Tools☆30Sep 24, 2024Updated last year
- scalable accelerated optimal control☆19Oct 8, 2022Updated 3 years ago
- Code for ICLR 2024 paper "When should we prefer Decision Transformers for Offline Reinforcement Learning?"☆17Jan 31, 2024Updated 2 years ago
- In Defense of the Unitary Scalarization for Deep Multi-Task Learning☆22Mar 8, 2023Updated 3 years ago
- A curated list of solvers/software/frameworksrelevant for dynamic optimiation☆36Jun 26, 2025Updated last year