Few-Shot Preference Optimization (FSPO) personalizes LLMs by reframing reward modeling as a meta-learning problem, enabling rapid adaptation to user preferences with minimal labeled data, leveraging synthetic datasets for scalability, and achieving high success rates in personalized content generation across multiple domains.
☆16Feb 27, 2025Updated last year
Alternatives and similar repositories for fewshot-preference-optimization
Users that are interested in fewshot-preference-optimization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The open-source repository for PAL: Sample-Efficient Personalized Reward Modeling for Pluralistic Alignment, which provides a general per…☆17Aug 28, 2025Updated 11 months ago
- ☆20Sep 16, 2025Updated 10 months ago
- Official implementation for Text Generation Beyond Discrete Token Sampling☆26Aug 11, 2025Updated 11 months ago
- Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"☆17Mar 31, 2025Updated last year
- All-in-one repository for Fine-tuning & Pretraining (Large) Language Models☆15Mar 8, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 登录脚本☆12Nov 4, 2022Updated 3 years ago
- The offical code for paper "What Constitutes a Faithful Summary? Preserving Author Perspectives in News Summarization"☆10Jun 23, 2024Updated 2 years ago
- True Few-Shot BioIE: Benchmarking GPT-3 In-Context and Small PLM Fine-Tuning☆12Jul 6, 2022Updated 4 years ago
- AbstainQA, ACL 2024☆29Feb 4, 2026Updated 6 months ago
- Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision☆124Sep 9, 2024Updated last year
- 无人机编队重构☆12Jul 28, 2018Updated 8 years ago
- 010Editor-Crack version:13.0.1☆10Mar 18, 2024Updated 2 years ago
- ☆28Jun 1, 2026Updated 2 months ago
- official implementation of ICLR'2025 paper: Rethinking Bradley-Terry Models in Preference-based Reward Modeling: Foundations, Theory, and…☆73Apr 2, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [KDD 2026] "Breaking Information Cocoons: A Hyperbolic Graph-LLM Framework for Exploration and Exploitation in Recommender Systems"☆16Jan 29, 2025Updated last year
- ☆10Jun 15, 2024Updated 2 years ago
- ☆15Nov 18, 2025Updated 8 months ago
- ☆12Jan 20, 2025Updated last year
- Apply Iprompt on GLM with innovative new methods. Currently support Chinese QA, English QA and Chinese poem generation.☆20Jun 16, 2022Updated 4 years ago
- Code base for "A General Contextualized Rewriting Framework for Text Summarization"☆13Jul 17, 2022Updated 4 years ago
- This is code for How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and Analysis☆18Nov 5, 2025Updated 9 months ago
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection☆12Feb 6, 2024Updated 2 years ago
- On-the-fly Definition Augmentation of LLMs for Biomedical NER☆14Apr 14, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Jan 20, 2024Updated 2 years ago
- Automatically installs and configures XFCE, XRDP and variables for a one-script setup☆14Apr 14, 2021Updated 5 years ago
- ☆16Dec 14, 2022Updated 3 years ago
- 60k hours of phoneme-aligned audio from audio books☆19Jul 27, 2024Updated 2 years ago
- Basic Tools☆13Dec 18, 2021Updated 4 years ago
- 中国科学院大学(国科大)研一课程☆19May 24, 2023Updated 3 years ago
- ☆12Mar 24, 2023Updated 3 years ago
- Code for the paper "Spectral Editing of Activations for Large Language Model Alignments"☆31Dec 20, 2024Updated last year
- GraphDancer: Training LLMs to Explore and Reason over Graphs via Curriculum Reinforcement Learning☆20May 25, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A simple Docker sandbox example and a ready-to-use autograder API. Based on asynchronous FastAPI and disposable Docker containers. Three …☆15Jan 10, 2022Updated 4 years ago
- [NeurIPS 2025] CAM: A Constructivist View of Agentic Memory for LLM-Based Reading Comprehension☆23Oct 8, 2025Updated 10 months ago
- ☆16Jul 10, 2022Updated 4 years ago
- Teacher - student distillation using DeepSpeed☆20Oct 7, 2022Updated 3 years ago
- The Official Repository of the Cryptonite Dataset☆23Feb 19, 2022Updated 4 years ago
- Code to reproduce results of our experiments using LoRe☆17Jun 10, 2026Updated last month
- Code repo for EMNLP 2023 paper "Auto-Instruct: Automatic Instruction Generation and Ranking for Black-Box Language Models"☆23Nov 13, 2023Updated 2 years ago