flint-xf-fan / Federated-RLHFLinks

[AAMAS 2025] Privacy-preserving and Personalized RLHF, with convergence guarantees. The Code contains experiments for training multiple instances of GPT-2 for personalized sentiment aligned text generation.
13Updated 8 months ago

Alternatives and similar repositories for Federated-RLHF

Users that are interested in Federated-RLHF are comparing it to the libraries listed below

Sorting: