Code for paper: Reward Uncertainty for Exploration in Preference-based Reinforcement Learning
☆15May 26, 2022Updated 4 years ago
Alternatives and similar repositories for rune
Users that are interested in rune are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pref-RL provides ready-to-use PbRL agents that are easily extensible.☆11Aug 31, 2022Updated 3 years ago
- ☆13Sep 24, 2024Updated last year
- Environment creation, pick and place example with custom gripper on H2017 robot arm☆15Mar 16, 2026Updated 4 months ago
- PyTorch implementations for Offline Preference-Based RL (PbRL) algorithms☆21Mar 24, 2025Updated last year
- Listwise Reward Estimation for Offline Preference-based Reinforcement Learning (ICML 2024)☆18Jun 18, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- The source code of the paper "Towards Problem of First Miss under Mobile EdgeCaching"☆11Apr 12, 2021Updated 5 years ago
- code for "Decoupled Preference-based Reinforcement Learning for Personalized Human-Robot Interaction"☆11Jul 9, 2022Updated 4 years ago
- ☆54Nov 10, 2022Updated 3 years ago
- Implementation of ICLR 2025 paper "Q-Adapter: Customizing Pre-trained LLMs to New Preferences with Forgetting Mitigation"☆18Oct 5, 2024Updated last year
- Inference API server with echo and gRPC to triton server (golang)☆13Nov 16, 2022Updated 3 years ago
- Context A real online retail transaction data set of two years. Content This Online Retail II data set contains all the transactions oc…☆18Jul 5, 2020Updated 6 years ago
- code for polite☆12Feb 28, 2024Updated 2 years ago
- [ICLR 2025] Highly Efficient Self-Adaptive Reward Shaping for Reinforcement Learning (SASR)☆12Aug 26, 2025Updated 11 months ago
- serialization benchmarks for common python libraries☆13Dec 31, 2018Updated 7 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 🍓 A toy object-oriented programming language written by rust☆17Apr 10, 2024Updated 2 years ago
- Single-Life Reinforcement Learning☆14Dec 17, 2022Updated 3 years ago
- Official code for ICML 2024 paper, "RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences" (ICML 2024 Spotlight)☆36Oct 15, 2024Updated last year
- ☆13Feb 5, 2024Updated 2 years ago
- Code for the content caching algorithm in edge caching.☆22Sep 24, 2024Updated last year
- Code for our ACL 2019 long paper: "Ensuring Readability and Data-fidelity using Head-modifier Templates in Deep Type Description Generati…☆11Nov 5, 2022Updated 3 years ago
- ☆14Jun 25, 2022Updated 4 years ago
- GPT-Critic: Offline Reinforcement Learning for End-to-End Task-Oriented Dialogue Systems☆10Jul 7, 2022Updated 4 years ago
- A Pytorch implementation of "Deep Learning with Logged Bandit Feedback"☆10Aug 22, 2018Updated 7 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Implementation and evaluation of Almanac (Automaton/Logic Multi-Agent Natural Actor-Critic), an algorithm for multi-agent reinforcement l…☆10May 5, 2022Updated 4 years ago
- TensorFlow implementation for our paper "Learning Long-Term Reward Redistribution via Randomized Return Decomposition"☆19Mar 17, 2022Updated 4 years ago
- ☆11Aug 10, 2020Updated 5 years ago
- ☆13Feb 5, 2025Updated last year
- ☆14Oct 11, 2022Updated 3 years ago
- Collapsed Gibbs sampling for Latent Dirichlet Allocation☆18Jun 11, 2012Updated 14 years ago
- We define and estimate smooth unique information of samples with respect to classifier weights and predictions. We compute these quantiti…☆11Mar 9, 2021Updated 5 years ago
- implementation of cooperative caching algorithm for edge computing☆17May 9, 2023Updated 3 years ago
- Code for AAAI 2021 long paper Learning from Crowds by Modeling Common Confusions.☆11Feb 6, 2021Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Model Primitive Hierarchical Reinforcement Learning☆13Dec 8, 2022Updated 3 years ago
- Deep learning based predictive analytics for efficient content caching in edge network☆18Dec 26, 2022Updated 3 years ago
- Benchmarks for occupancy mapping libraries in Robotics☆21Mar 17, 2019Updated 7 years ago
- Policy Transfer across Visual and Dynamics Domain Gaps via Iterative Grounding (RSS 2021)☆12Oct 22, 2021Updated 4 years ago
- Classification of animal sounds in a hyperdiverse rainforest using Convolutional Neural Networks (Sun et al, 2021)☆13Oct 16, 2023Updated 2 years ago
- Attempted implementation of a Bi-directional GRU followed by a linear-chain-CRF (from scratch) for Named Entity Recognition.☆15Dec 5, 2017Updated 8 years ago
- ☆26Feb 19, 2024Updated 2 years ago