flint-xf-fan / Federated-RLHFView on GitHub
[AAMAS 2025] Privacy-preserving and Personalized RLHF, with convergence guarantees. The Code contains experiments for training multiple instances of GPT-2 for personalized sentiment aligned text generation.
15Apr 16, 2025Updated 10 months ago

Alternatives and similar repositories for Federated-RLHF

Users that are interested in Federated-RLHF are comparing it to the libraries listed below

Sorting:

Are these results useful?