Official code for ICML 2024 paper, "RIME: Robust Preference-based Reinforcement Learning with Noisy Preferences" (ICML 2024 Spotlight)
☆36Oct 15, 2024Updated last year
Alternatives and similar repositories for RIME_ICML2024
Users that are interested in RIME_ICML2024 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official code for the ICLR 2025 paper, "Scaling Offline Model-Based RL via Jointly-Optimized World-Action Model Pretraining"☆30Dec 1, 2024Updated last year
- ☆13Sep 24, 2024Updated 2 years ago
- Listwise Reward Estimation for Offline Preference-based Reinforcement Learning (ICML 2024)☆18Jun 18, 2024Updated 2 years ago
- This is the official code repository for the paper "Decoding Global Preferences: Temporal and Cooperative Dependency Modeling in Multi-Ag…☆11Apr 9, 2026Updated 5 months ago
- Official code for the paper, "Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning"☆174Oct 23, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for paper: Reward Uncertainty for Exploration in Preference-based Reinforcement Learning☆15May 26, 2022Updated 4 years ago
- PyTorch implementations for Offline Preference-Based RL (PbRL) algorithms☆23Mar 24, 2025Updated last year
- Official Codebase for TMLR 2023, Benchmarks and Algorithms for Offline Preference-Based Reward Learning☆20Dec 30, 2022Updated 3 years ago
- ☆12Feb 21, 2025Updated last year
- ☆25May 20, 2025Updated last year
- code for polite☆12Feb 28, 2024Updated 2 years ago
- ICCV 2023 - AdaptGuard: Defending Against Universal Attacks for Model Adaptation☆11Dec 23, 2023Updated 2 years ago
- Scaling Population-Based Reinforcement Learning with GPU Accelerated Simulation☆13Nov 5, 2025Updated 10 months ago
- [ICRA'25] NeuGrasp: Generalizable Neural Surface Reconstruction with Background Priors for Material-Agnostic Object Grasp Detection☆22Jan 29, 2026Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆38Apr 27, 2023Updated 3 years ago
- Uni-RLHF platform for "Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback" (ICLR2024…☆42Nov 20, 2024Updated last year
- ☆13Sep 11, 2026Updated 2 weeks ago
- calibrate EIDM model in SUMO using HighD data set☆16May 9, 2025Updated last year
- Code and dataset for the ICLR 2024 paper "Thought Propagation: An analogical Approach to Complex Reasoning with Large Language Models."☆17Mar 4, 2024Updated 2 years ago
- [ICLR 2024] Adaptive Replay Ratio implementation from 'Revisiting Plasticity in Visual RL: Data, Modules and Training Stages'.☆13Oct 9, 2024Updated last year
- ☆43May 25, 2023Updated 3 years ago
- Adaptive explicit-Barrier Net for Safe and Scalable Robot Learning☆19May 16, 2025Updated last year
- ☆11May 29, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This repository contains Reinforcement Learning (RL) environments for the Upkie robot.☆36May 26, 2026Updated 4 months ago
- [IROS 2024] PP-TIL: Personalized Planning for Autonomous Driving with Instance-based Transfer Imitation Learning☆24Feb 28, 2025Updated last year
- Official code for the long-horizon language-conditioned robotic manipulation benchmark LoHoRavens.☆21Oct 8, 2024Updated last year
- Tensorflow implementation of the TGRS paper entitled Oil Spill Segmentation via Adversarial f-Divergence Learning.☆12Mar 9, 2019Updated 7 years ago
- RLHF-Blender: A Configurable Interactive Interface for Learning from Diverse Human Feedback☆14Sep 18, 2026Updated last week
- https://arxiv.org/abs/2312.10807☆85Jun 22, 2026Updated 3 months ago
- [ICML 2024] Learning Reward for Robot Skills Using Large Language Models via Self-Alignment☆19Aug 22, 2024Updated 2 years ago
- [ICML2026] Official JAX code for Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying☆17Jul 3, 2026Updated 2 months ago
- Code to reproduce results from the paper: Prediction and Control in Continual Reinforcement Learning, NeurIPS 2023.☆13May 10, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR'26] Official Repository for The Paper: Stop Tracking Me! Proactive Defense Against Attribute Inference Attack in LLMs☆26Apr 6, 2026Updated 5 months ago
- ☆17Jul 25, 2023Updated 3 years ago
- ☆22Apr 17, 2026Updated 5 months ago
- OpenAI-gym-like Reinforcement Learning environment for Dispatching of Mobile Chargers with SUMO. Compatible with Gym and popular RL libra…☆15Mar 16, 2025Updated last year
- ☆22Dec 17, 2020Updated 5 years ago
- Central repository for all baxter working code.☆21Jul 18, 2025Updated last year
- ☆10May 10, 2024Updated 2 years ago