Few-Shot Preference Optimization (FSPO) personalizes LLMs by reframing reward modeling as a meta-learning problem, enabling rapid adaptation to user preferences with minimal labeled data, leveraging synthetic datasets for scalability, and achieving high success rates in personalized content generation across multiple domains.
☆17Feb 27, 2025Updated last year
Alternatives and similar repositories for fewshot-preference-optimization
Users that are interested in fewshot-preference-optimization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The open-source repository for PAL: Sample-Efficient Personalized Reward Modeling for Pluralistic Alignment, which provides a general per…☆17Aug 28, 2025Updated last year
- ☆20Sep 16, 2025Updated last year
- ☆19Oct 8, 2024Updated 2 years ago
- Official implementation for Text Generation Beyond Discrete Token Sampling☆26Aug 11, 2025Updated last year
- Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"☆17Mar 31, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- All-in-one repository for Fine-tuning & Pretraining (Large) Language Models☆15Mar 8, 2023Updated 3 years ago
- Speechflow for emotion recognition related information decomposition☆10Jul 27, 2021Updated 5 years ago
- The offical code for paper "What Constitutes a Faithful Summary? Preserving Author Perspectives in News Summarization"☆10Jun 23, 2024Updated 2 years ago
- True Few-Shot BioIE: Benchmarking GPT-3 In-Context and Small PLM Fine-Tuning☆12Jul 6, 2022Updated 4 years ago
- AbstainQA, ACL 2024☆30Feb 4, 2026Updated 8 months ago
- 无人机编队重构☆12Jul 28, 2018Updated 8 years ago
- 010Editor-Crack version:13.0.1☆10Mar 18, 2024Updated 2 years ago
- ☆28Jun 1, 2026Updated 4 months ago
- official implementation of ICLR'2025 paper: Rethinking Bradley-Terry Models in Preference-based Reward Modeling: Foundations, Theory, and…☆73Apr 2, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆10Jun 15, 2024Updated 2 years ago
- [KDD 2026 Oral] "Breaking Information Cocoons: A Hyperbolic Graph-LLM Framework for Exploration and Exploitation in Recommender Systems"☆18Jan 29, 2025Updated last year
- 一个简易的8层电梯控制器,使用verilog HDL语言描述 / a simple elevator controller works with verilog HDL☆12Jul 15, 2020Updated 6 years ago
- ☆15Nov 18, 2025Updated 10 months ago
- ☆12Jan 20, 2025Updated last year
- Official Implementation of "Personalized Pieces: Efficient Personalized Large Language Models through Collaborative Efforts" at EMNLP 202…☆13Oct 27, 2024Updated last year
- Apply Iprompt on GLM with innovative new methods. Currently support Chinese QA, English QA and Chinese poem generation.☆20Jun 16, 2022Updated 4 years ago
- Code base for "A General Contextualized Rewriting Framework for Text Summarization"☆13Jul 17, 2022Updated 4 years ago
- Official code for our paper "Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models"☆25Oct 31, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is code for How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and Analysis☆18Updated this week
- On-the-fly Definition Augmentation of LLMs for Biomedical NER☆14Apr 14, 2025Updated last year
- ☆12Jul 6, 2023Updated 3 years ago
- ☆14Sep 7, 2022Updated 4 years ago
- ☆12Jan 20, 2024Updated 2 years ago
- Automatically installs and configures XFCE, XRDP and variables for a one-script setup☆14Apr 14, 2021Updated 5 years ago
- Basic Tools☆13Dec 18, 2021Updated 4 years ago
- 中国科学院大学(国科大)研一课程☆20May 24, 2023Updated 3 years ago
- ☆12Mar 24, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Spectral Editing of Activations for Large Language Model Alignments"☆31Dec 20, 2024Updated last year
- GraphDancer: Training LLMs to Explore and Reason over Graphs via Curriculum Reinforcement Learning☆21May 25, 2026Updated 4 months ago
- ☆17Oct 30, 2022Updated 3 years ago
- A simple Docker sandbox example and a ready-to-use autograder API. Based on asynchronous FastAPI and disposable Docker containers. Three …☆15Jan 10, 2022Updated 4 years ago
- How well can Text-to-Image Generative Models understand Ethical Natural Language Interventions?☆12Aug 16, 2023Updated 3 years ago
- tacotronV2 + wavernn 实现中文语音合成(Tensorflow + pytorch)☆15May 20, 2020Updated 6 years ago
- 智慧树刷互动分,自动复读问答☆10Nov 19, 2020Updated 5 years ago