Few-Shot Preference Optimization (FSPO) personalizes LLMs by reframing reward modeling as a meta-learning problem, enabling rapid adaptation to user preferences with minimal labeled data, leveraging synthetic datasets for scalability, and achieving high success rates in personalized content generation across multiple domains.
☆17Feb 27, 2025Updated last year
Alternatives and similar repositories for fewshot-preference-optimization
Users that are interested in fewshot-preference-optimization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Sep 16, 2025Updated last year
- ☆19Oct 8, 2024Updated last year
- Official implementation for Text Generation Beyond Discrete Token Sampling☆26Aug 11, 2025Updated last year
- Code for reproducing our paper "Low Rank Adapting Models for Sparse Autoencoder Features"☆17Mar 31, 2025Updated last year
- All-in-one repository for Fine-tuning & Pretraining (Large) Language Models☆15Mar 8, 2023Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- 登录脚本☆12Nov 4, 2022Updated 3 years ago
- The offical code for paper "What Constitutes a Faithful Summary? Preserving Author Perspectives in News Summarization"☆10Jun 23, 2024Updated 2 years ago
- True Few-Shot BioIE: Benchmarking GPT-3 In-Context and Small PLM Fine-Tuning☆12Jul 6, 2022Updated 4 years ago
- AbstainQA, ACL 2024☆30Feb 4, 2026Updated 7 months ago
- Easy-to-Hard Generalization: Scalable Alignment Beyond Human Supervision☆124Sep 9, 2024Updated 2 years ago
- 010Editor-Crack version:13.0.1☆10Mar 18, 2024Updated 2 years ago
- SECOM: On Memory Construction and Retrieval for Personalized Conversational Agents, ICLR 2025☆63Mar 1, 2025Updated last year
- [KDD 2026] "Breaking Information Cocoons: A Hyperbolic Graph-LLM Framework for Exploration and Exploitation in Recommender Systems"☆16Jan 29, 2025Updated last year
- ☆10Jun 15, 2024Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- ☆15Nov 18, 2025Updated 10 months ago
- Official Implementation of "Personalized Pieces: Efficient Personalized Large Language Models through Collaborative Efforts" at EMNLP 202…☆13Oct 27, 2024Updated last year
- Apply Iprompt on GLM with innovative new methods. Currently support Chinese QA, English QA and Chinese poem generation.☆20Jun 16, 2022Updated 4 years ago
- ☆13Apr 17, 2018Updated 8 years ago
- Code base for "A General Contextualized Rewriting Framework for Text Summarization"☆13Jul 17, 2022Updated 4 years ago
- Official code for our paper "Reasoning Models Hallucinate More: Factuality-Aware Reinforcement Learning for Large Reasoning Models"☆25Oct 31, 2025Updated 10 months ago
- [ICRA 2024] WLST: Weak Labels Guided Self-training for Weakly-supervised Domain Adaptation on 3D Object Detection☆12Feb 6, 2024Updated 2 years ago
- On-the-fly Definition Augmentation of LLMs for Biomedical NER☆14Apr 14, 2025Updated last year
- ☆12Jul 6, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆14Sep 7, 2022Updated 4 years ago
- this is based on the paper Chain-of-Retrieval Augmented Generation☆15Mar 29, 2025Updated last year
- ☆12Jan 20, 2024Updated 2 years ago
- ☆16Dec 14, 2022Updated 3 years ago
- [NeurIPS 2024] Train LLMs with diverse system messages reflecting individualized preferences to generalize to unseen system messages☆53Aug 10, 2025Updated last year
- ☆20Feb 17, 2024Updated 2 years ago
- Basic Tools☆13Dec 18, 2021Updated 4 years ago
- Making of cuda kernel☆17May 27, 2025Updated last year
- 中国科学院大学(国科大)研一课程☆20May 24, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for the paper "Spectral Editing of Activations for Large Language Model Alignments"☆31Dec 20, 2024Updated last year
- GraphDancer: Training LLMs to Explore and Reason over Graphs via Curriculum Reinforcement Learning☆21May 25, 2026Updated 3 months ago
- A simple Docker sandbox example and a ready-to-use autograder API. Based on asynchronous FastAPI and disposable Docker containers. Three …☆15Jan 10, 2022Updated 4 years ago
- How well can Text-to-Image Generative Models understand Ethical Natural Language Interventions?☆13Aug 16, 2023Updated 3 years ago
- 智慧树刷互动分,自动复读问答☆10Nov 19, 2020Updated 5 years ago
- [NeurIPS 2025] CAM: A Constructivist View of Agentic Memory for LLM-Based Reading Comprehension☆22Oct 8, 2025Updated 11 months ago
- Official implementation of paper "Learning High-Order Relationships of Brain Regions" [ICML 2024]☆24May 29, 2025Updated last year