The Image as Its Own Reward: Reinforcement Learning with Adversarial Reward for Image Generation
图像作为自身奖励:基于对抗性奖励的图像生成强化学习
机构 * Show Lab, National University of Singapore(新加坡国立大学Show实验室)
专题命中 图像生成评测 :image generation(title,abstract);分类 cs.CV
AI总结 本文提出Adv-GRPO框架,通过对抗性奖励提升图像生成质量,利用图像自身作为奖励,结合参考图像和基础模型,实现更高质量和审美效果。