Training-Free Self-Correction for Multimodal Masked Diffusion Models
无需训练的多模态掩码扩散模型自校正
机构 * University of California, Los Angeles(加州大学洛杉矶分校) ; Mohamed bin Zayed University of Artificial Intelligence(莫莫德·本·扎耶德人工智能大学) ; East China Normal University(华东师范大学) ; University of Virginia(弗吉尼亚大学) ; University of Minnesota(明尼苏达大学) ; Drexel university(德雷塞尔大学) ; The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) ; University of Toronto(多伦多大学)
专题命中 多模态生成 :multimodal(title,abstract)
AI总结 本文提出无需训练的多模态掩码扩散模型自校正方法,通过减少采样步骤提升生成质量,适用于文本到图像和多模态理解任务。