Kwai-Kolors / LPOLinks

Diffusion Model as a Noise-Aware Latent Reward Model for Step-Level Preference Optimization
20Updated last month

Alternatives and similar repositories for LPO

Users that are interested in LPO are comparing it to the libraries listed below

Sorting: