Diffusion-SDPO: Safeguarded Direct Preference Optimization for Diffusion Models
Diffusion-SDPO:用于扩散模型的受保护直接偏好优化
机构 * School of Artificial Intelligence, Nanjing University(人工智能学院,南京大学) ; National Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家实验室,南京大学) ; Alibaba International Digital Commerce Group(阿里巴巴国际数字商业集团)
专题命中 病理影像 :pathology(abstract);分类 cs.CV
AI总结 Diffusion-SDPO通过保护性更新规则提升扩散模型对人类偏好的对齐效果。
Comments The code is publicly available at https://github.com/AIDC-AI/Diffusion-SDPO