Relative Preference Optimization: Enhancing LLM Alignment through Contrasting Responses across Identical and Diverse Prompts
☆26Feb 23, 2024Updated 2 years ago
Alternatives and similar repositories for relative-preference-optimization
Users that are interested in relative-preference-optimization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- ☆18Aug 4, 2025Updated last year
- Official implementation of Bootstrapping Language Models via DPO Implicit Rewards☆49Apr 15, 2025Updated last year
- Gamma Belief Networks☆12Mar 2, 2018Updated 8 years ago
- Extensive Self-Contrast Enables Feedback-Free Language Model Alignment☆20Apr 2, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆23Oct 10, 2025Updated 10 months ago
- 这是我的博客《不用框架,使用Python搭建基于numpy的卷积神经网络来进行cifar-10分类的深度学习系统》的代码实现。☆10Jul 1, 2019Updated 7 years ago
- DuoGuard: A Two-Player RL-Driven Framework for Multilingual LLM Guardrails☆34Feb 26, 2025Updated last year
- ☆45Sep 19, 2024Updated last year
- ☆13Jul 2, 2025Updated last year
- ☆10Oct 15, 2019Updated 6 years ago
- ☆12Jan 4, 2024Updated 2 years ago
- Latest Evaluation Toolkit (LatestEval). Assessing the language models with latest, uncontaminated materials.☆29Feb 17, 2025Updated last year
- Dual-level Adaptive Self-Labeling for Novel Class Discovery in Point Cloud Segmentation (ECCV2024)☆15Nov 1, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PyTorch implementation of CARE☆16Oct 6, 2023Updated 2 years ago
- ☆17Jul 25, 2023Updated 3 years ago
- Reproduction of "RLCD Reinforcement Learning from Contrast Distillation for Language Model Alignment☆70Aug 18, 2023Updated 3 years ago
- The code and data for the paper JiuZhang3.0☆49May 26, 2024Updated 2 years ago
- ☆19Jul 30, 2025Updated last year
- 从零开始无框架python实现卷积神经网络☆13Aug 24, 2020Updated 6 years ago
- Official code for our COLING 2022 paper: In-Context Learning for Empathetic Dialogue Generation☆20Mar 1, 2023Updated 3 years ago
- Explore, Establish, Exploit: Red Teaming Language Models from Scratch☆15Jun 21, 2023Updated 3 years ago
- A Multi-task learning framework for personality trait detection☆11Jan 13, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Overcooked-AI Experiment Psiturk Demo (for MTurk experiments)☆13May 10, 2021Updated 5 years ago
- Official eval scripts for JobBench☆49Updated this week
- ICCV'23 | Adverse Weather Removal with Codebook Priors☆10Aug 28, 2023Updated 3 years ago
- ☆30Dec 27, 2024Updated last year
- ☆22Jan 17, 2025Updated last year
- Code for ACL2024 paper - Adversarial Preference Optimization (APO).☆55Jun 3, 2024Updated 2 years ago
- AI for Business is Now, and the Future.☆12Oct 28, 2023Updated 2 years ago
- ☆21May 16, 2024Updated 2 years ago
- Source codes for "Preference-grounded Token-level Guidance for Language Model Fine-tuning" (NeurIPS 2023).☆17Jan 8, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆19Oct 2, 2023Updated 2 years ago
- The PyTorch code for paper: "CONSK-GCN: Conversational Semantic- and Knowledge-Oriented Graph Convolutional Network for Multimodal Emotio…☆13Oct 21, 2022Updated 3 years ago
- ☆10Jul 4, 2024Updated 2 years ago
- ☆19Oct 8, 2024Updated last year
- A toolkit for synthesizing high-quality code training data using LLM agents. It provides three independent pipelines, each producing a di…☆17Mar 10, 2026Updated 5 months ago
- [NeurIPS'23] Binary Classification with Confidence Difference☆10May 13, 2024Updated 2 years ago
- Code and models for EMNLP 2024 paper "WPO: Enhancing RLHF with Weighted Preference Optimization"☆41Sep 24, 2024Updated last year