Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"
☆35Dec 14, 2023Updated 2 years ago
Alternatives and similar repositories for hidden-context
Users that are interested in hidden-context are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Dec 17, 2020Updated 5 years ago
- ☆16Apr 12, 2023Updated 3 years ago
- ☆48Mar 25, 2025Updated last year
- Reproduction of OpenAI and DeepMind's "Deep Reinforcement Learning from Human Preferences"☆31Jul 27, 2021Updated 4 years ago
- Code for the paper "Learning Temporal Distances: Contrastive Successor Features Can Provide a Metric Structure for Decision-Making"☆29Jul 11, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Rewarded soups official implementation☆64Sep 27, 2023Updated 2 years ago
- Code for Model-Free Opponent Shaping (ICML 2022)☆24Nov 18, 2022Updated 3 years ago
- Source code of "Variational Imitation Learning with Diverse-quality Demonstrations" in ICML 2020. This github repository includes python …☆20Aug 16, 2021Updated 4 years ago
- A curated reading list for researchers in the Philosophy of Interpretability☆17Aug 17, 2025Updated 11 months ago
- ☆164Nov 23, 2024Updated last year
- Baselines for Neural MMO -- new users should treat this repo as a starter project☆52Jul 29, 2024Updated last year
- ☆15Jun 17, 2025Updated last year
- PyTorch implementation of "The Option Keyboard: Combining Skills in Reinforcement Learning" (NeurIPS 2019)☆12Jul 2, 2020Updated 6 years ago
- code to reproduce the empirical results in the research paper☆40Oct 12, 2021Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for EMNLP'24 paper - On Diversified Preferences of Large Language Model Alignment☆16Aug 6, 2024Updated last year
- Code and data for the paper "Bridging RL Theory and Practice with the Effective Horizon"☆50Jun 26, 2024Updated 2 years ago
- This repo support auto line plot for multi-seed event file from TensorBoard☆12Jun 23, 2022Updated 4 years ago
- UCLA CS 188 (Winter 2023) course project.☆12Mar 31, 2023Updated 3 years ago
- ☆25Jun 13, 2024Updated 2 years ago
- TensorFlow implementation for our paper "Learning Long-Term Reward Redistribution via Randomized Return Decomposition"☆19Mar 17, 2022Updated 4 years ago
- Decoding of the speech envelope from EEG using the VLAAI deep neural network☆14Sep 28, 2022Updated 3 years ago
- Code to reproduce the experiments in The Mirage of Action-Dependent Baselines in Reinforcement Learning.☆17Aug 2, 2018Updated 7 years ago
- Implementing REINFORCE algorithm on Pong, Lunar Lander and Cartplot + Medium Article☆23Nov 24, 2020Updated 5 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ACL'24, Outstanding Paper] Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!☆39Aug 2, 2024Updated last year
- This repo is reproduction resources for linear alignment paper, still working☆17May 19, 2024Updated 2 years ago
- Source code for Interpretable Reward Redistribution in Reinforcement Learning: A Causal Approach (NeurIPS 2023)☆10Dec 12, 2023Updated 2 years ago
- OpenLLMDE: An open source data engineering framework for LLMs☆18Sep 9, 2023Updated 2 years ago
- K* search based implementation of top-k and top-quality planners☆19Apr 1, 2026Updated 3 months ago
- Risk-sensitive Inverse Reinforcement Learning☆11Sep 11, 2019Updated 6 years ago
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 4 months ago
- Robust Reinforcement Learning Benchmark☆13Sep 22, 2024Updated last year
- Python neighbor-joining library. Goal: Efficient O(n^2) neighbor-joining algorithm.☆12May 5, 2014Updated 12 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Code for the paper Novelty Search in Representational Space for Sample Efficient Exploration presented at NeurIPS 2020.☆14Jul 16, 2024Updated 2 years ago
- Adversaial attack comparative assessment Large Language Model☆13May 21, 2025Updated last year
- [TMLR 2025] A collection of research papers on constraint inference within the field of RL☆11May 9, 2025Updated last year
- ☆12Jul 19, 2022Updated 4 years ago
- Code for the ICML 2021 paper "Bridging Multi-Task Learning and Meta-Learning: Towards Efficient Training and Effective Adaptation", Haoxi…☆67Oct 18, 2021Updated 4 years ago
- DINOv2 module for use with Autodistill.☆16Dec 6, 2023Updated 2 years ago
- This is the source code of RPG (Reward-Randomized Policy Gradient)☆42Sep 1, 2022Updated 3 years ago