[AAMAS 2025] Privacy-preserving and Personalized RLHF, with convergence guarantees. The Code contains experiments for training multiple instances of GPT-2 for personalized sentiment aligned text generation.
rft federated-reinforcement-learning llms rlhf reinforcement-learning-from-human-feedback fedrl personalized-rlhf fedrlhf federated-rlhf
-
Updated
Apr 5, 2025 - Python