mihirp1998

mihirp1998/AlignProp

AlignProp uses direct reward backpropogation for the alignment of large-scale text-to-image diffusion models. Our method is 25x more sample and compute efficient than reinforcement learning methods (PPO) for finetuning Stable Diffusion

⭐ 326 ⑂ 11 Python MIT · 2024-11-01推送
326
Watchers
0
贡献者
0
Commits
0
Releases
6
Open Issues
2024-11-01
最近推送
原文 中文
暂无 README