arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-01-15 至 2026-01-15 共收录 3 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3 篇

2309.06135 2026-01-15 cs.CL cs.CV 88%

Prompting4Debugging: Red-Teaming Text-to-Image Diffusion Models by Finding Problematic Prompts

Prompting4Debugging: 通过寻找问题性提示对文本到图像扩散模型进行红队测试

Zhi-Yi Chin, Chieh-Ming Jiang, Ching-Chun Huang, Pin-Yu Chen, Wei-Chen Chiu

机构 * Department of Computer Science, National Yang Ming Chiao Tung University, Hsinchu, Taiwan(国家阳明交通大学计算机科学系) IBM Research, NY 10598, USA(IBM研究院)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

AI总结 Prompting4Debugging通过寻找问题性提示揭示文本到图像扩散模型的安全漏洞,表明现有安全机制存在重大缺陷。

Comments ICML 2024 main conference paper. The source code is available at https://github.com/zhiyichin/P4D

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11520 2026-01-15 cs.CV 88%

LaCon: Late-Constraint Diffusion for Steerable Guided Image Synthesis

LaCon:晚期约束扩散用于可操控引导图像合成

Chang Liu, Rui Li, Kaidong Zhang, Xin Luo, Dong Liu

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 文生图 :diffusion(title,abstract);image synthesis(title,abstract);分类 cs.CV

AI总结 LaCon通过在预训练扩散模型中整合多种条件,实现灵活可控的图像合成,提升了扩散模型的泛化能力和效率。

Comments GitHub repo: https://github.com/AlonzoLeeeooo/LCDG

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09647 2026-01-15 cs.CV cs.CR cs.LG 79%

Identifying Models Behind Text-to-Image Leaderboards

识别文本到图像排行榜背后的模型

Ali Naseh, Yuefeng Peng, Anshuman Suri, Harsh Chaudhari, Alina Oprea, Amir Houmansadr

机构 * University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校) Northeastern University(东北大学)

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV

AI总结 研究揭示了文本到图像排行榜中通过图像嵌入空间聚类实现模型匿名性的突破,发现模型特定特征及提示对可区分性的影响,揭示了排行榜中的安全漏洞。

详情

展开后加载摘要…

URL PDF HTML 收藏