mihirp1998/AlignProp
AlignProp uses direct reward backpropogation for the alignment of large-scale text-to-image diffusion models. Our method is 25x more sample and compute efficient than reinforcement learning methods (PPO) for finetuning Stable Diffusion
⭐ 326
⑂ 11
Python
MIT
· 2024-11-01推送
326
Watchers
0
贡献者
0
Commits
0
Releases
6
Open Issues
2024-11-01
最近推送
原文
中文
暂无 README