flint-xf-fan / Federated-RLHF
View external linksLinks

[AAMAS 2025] Privacy-preserving and Personalized RLHF, with convergence guarantees. The Code contains experiments for training multiple instances of GPT-2 for personalized sentiment aligned text generation.
14Apr 16, 2025Updated 10 months ago

Alternatives and similar repositories for Federated-RLHF

Users that are interested in Federated-RLHF are comparing it to the libraries listed below

Sorting:

Are these results useful?