arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-03-05 至 2026-03-05 共收录 3 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3 篇

2603.02829 2026-03-05 cs.CV cs.LG 88%

Toward Early Quality Assessment of Text-to-Image Diffusion Models

面向文本到图像扩散模型早期质量评估

Huanlei Guo, Hongxin Wei, Bingyi Jing

机构 * Southern University of Science and Technology(南方科技大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Shenzhen Loop Area Institute(深圳河套学院)

专题命中 文生图 :text-to-image(title,abstract);diffusion(title,abstract);分类 cs.CV

AI总结 本研究提出Probe-Select模块,通过早期评估指导文本到图像扩散模型的生成过程,降低采样成本并提升保留图像质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20376 2026-03-05 cs.CV cs.CR 88%

When Memory Becomes a Vulnerability: Towards Multi-turn Jailbreak Attacks against Text-to-Image Generation Systems

当记忆成为漏洞:针对文本到图像生成系统的多轮劫持攻击

Shiqian Zhao, Jiayang Liu, Yiming Li, Runyi Hu, Xiaojun Jia, Wenshu Fan, Xiao Bao, Xinfeng Li, Jie Zhang, Wei Dong, Tianwei Zhang, Luu Anh Tuan

机构 * Nanyang Technological University(南洋理工大学) Institute of Science(科学研究院) University of Electronic Science and Technology of China(电子科技大学) CFAR and IHPC Agency for Science Technology and Research(CFAR和IHPC科技研究局) VinUniversity(文大学)

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

AI总结 本文提出Inception,首个利用文本到图像生成系统内存机制的多轮劫持攻击方法,通过分割和递归模块有效规避安全过滤,提升攻击成功率20%。

Comments This work proposes a multi-turn jailbreak attack against real-world chat-based T2I generation systems that intergrate memory mechanism. It also constructed a simulation system, with considering three industrial-grade memory mechanisms, 7 kinds of safety filters (both input and output); It is going to appear on USENIX 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04267 2026-03-05 cs.CV 70%

UniLight: A Unified Representation for Lighting

UniLight: 一种统一的光照表示

Zitian Zhang, Iliyan Georgiev, Michael Fischer, Yannick Hold-Geoffroy, Jean-François Lalonde, Valentin Deschaintre

机构 * Université Laval(拉瓦尔大学) Adobe Research(Adobe研究)

专题命中 文生图 :diffusion(abstract);image synthesis(abstract);分类 cs.CV

AI总结 UniLight通过联合潜在空间统一多种光照表示,实现跨模态的光照特征提取与迁移,支持光照检索、环境映射生成和扩散模型中的光照控制。

Comments Project page: https://lvsn.github.io/UniLight

详情

展开后加载摘要…

URL PDF HTML 收藏