☆54Nov 10, 2022Updated 3 years ago
Alternatives and similar repositories for B_Pref
Users that are interested in B_Pref are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official codebase for "B-Pref: Benchmarking Preference-BasedReinforcement Learning" contains scripts to reproduce experiments.☆136Nov 3, 2021Updated 4 years ago
- ☆38Apr 27, 2023Updated 3 years ago
- ☆17Oct 12, 2023Updated 2 years ago
- Code for paper: Reward Uncertainty for Exploration in Preference-based Reinforcement Learning☆15May 26, 2022Updated 4 years ago
- ☆13Feb 5, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Preference Transformer: Modeling Human Preferences using Transformers for RL (ICLR2023 Accepted)☆168Oct 15, 2023Updated 2 years ago
- Official code for "Pretraining Representations For Data-Efficient Reinforcement Learning" (NeurIPS 2021)☆56Jul 27, 2021Updated 4 years ago
- Guide Your Agent with Adaptive Multimodal Rewards (NeurIPS 2023 Accepted)☆33Sep 25, 2023Updated 2 years ago
- Evaluating different engineering tricks that make RL work☆15Jun 3, 2021Updated 5 years ago
- ☆10Oct 11, 2022Updated 3 years ago
- Codes for Evolving Plastic ANNs☆15Dec 18, 2022Updated 3 years ago
- Offline RLHF codebase implementation for "Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human …☆42Mar 26, 2024Updated 2 years ago
- Environment creation, pick and place example with custom gripper on H2017 robot arm☆15Mar 16, 2026Updated 4 months ago
- Public code release for the paper "Reawakening knowledge: Anticipatory recovery from catastrophic interference via structured training"☆11Oct 27, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Jaehyung Kim et al's ACL 2023 paper on "infoVerse: A Universal Framework for Dataset Characterization with Multidimensional Meta-informat…☆16Jun 28, 2023Updated 3 years ago
- Official codebase for Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings.☆21Mar 5, 2021Updated 5 years ago
- ☆18Jun 8, 2023Updated 3 years ago
- Dataset collection and training code for "Ask Your Humans: Using Human Instructions to Improve Generalization in Reinforcement Learning"☆11Apr 8, 2025Updated last year
- A Library for Active Preference-based Reward Learning Algorithms☆55Dec 16, 2023Updated 2 years ago
- Trajectory-ranked Reward EXtrapolation (T-REX) for Inverse Reinforcement Learning - A Tensorflow implementation trained on OpenAI Gym env…☆19Jul 4, 2019Updated 7 years ago
- ☆16Apr 14, 2026Updated 3 months ago
- PyTorch implementations for Offline Preference-Based RL (PbRL) algorithms☆21Mar 24, 2025Updated last year
- [NeurIPS'21] RoMA: Robust Model Adaptation for Offline Model-based Optimization☆15Oct 28, 2021Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Latent Dynamics Mixture, NeurIPS 2021☆18Oct 25, 2022Updated 3 years ago
- This is my Curriculum Vitae☆14Mar 19, 2026Updated 4 months ago
- Simulation environments for Multi-Objective Reinforcement Learning (MORL)☆17Aug 2, 2022Updated 3 years ago
- Reproduction of OpenAI and DeepMind's "Deep Reinforcement Learning from Human Preferences"☆337Nov 29, 2021Updated 4 years ago
- ☆10Oct 3, 2023Updated 2 years ago
- ☆61Apr 16, 2023Updated 3 years ago
- Evaluating Safety of Autonomous Agents in Mobile Device Control (AAAI 2026 AI Alignment Track)☆34Jan 28, 2026Updated 5 months ago
- ☆15Aug 9, 2021Updated 4 years ago
- The official repository of Decoupled Reinforcement Learning to Stabilise Intrinsically-Motivated Exploration" (AAMAS 2022)☆26Feb 3, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Pre-Trained Language Models for Interactive Decision-Making [NeurIPS 2022]☆131Jun 8, 2022Updated 4 years ago
- Code for 'Mapping State Space using Landmarks for Universal Goal Reaching'.☆16Dec 26, 2023Updated 2 years ago
- Trajectory-wise Multiple Choice Learning for Dynamics Generalization in Reinforcement Learning (NeurIPS 2020)☆39Oct 27, 2020Updated 5 years ago
- Code for "World Model as a Graph: Learning Latent Landmarks for Planning" (ICML 2021 Long Presentation)☆71Jul 17, 2021Updated 5 years ago
- ☆61Feb 3, 2023Updated 3 years ago
- Visual Representation Learning with Stochastic Frame Prediction (ICML 2024)☆28Nov 27, 2024Updated last year
- Learning Invariant Representations for Reinforcement Learning without Reconstruction☆157Aug 31, 2021Updated 4 years ago