A Simple and Effective Reinforcement Learning Method for Text-to-Image Diffusion Fine-tuning
一种简单而有效的文本到图像扩散微调强化学习方法
机构 * University of Amsterdam(阿姆斯特丹大学) ; Meta ; Radboud University(拉德堡德大学)
AI总结 本文提出LOOP方法,结合REINFORCE的方差减少技术和PPO的鲁棒性,提升扩散模型在黑盒目标上的样本效率和性能。
Comments Published at Transactions on Machine Learning Research (TMLR), 2026