zhaoyl18 / SEIKOView on GitHub
SEIKO is a novel reinforcement learning method to efficiently fine-tune diffusion models in an online setting. Our methods outperform all baselines (PPO, classifier-based guidance, direct reward backpropagation) for fine-tuning Stable Diffusion.
30Jul 18, 2024Updated 2 years ago

Alternatives and similar repositories for SEIKO

Users that are interested in SEIKO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

Are these results useful?