arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-31 至 2025-10-31 共收录 3 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3 篇

2510.26105 2025-10-31 cs.CV cs.AI cs.CR 77%

Security Risk of Misalignment between Text and Image in Multi-modal Model

Xiaosen Wang, Zhijin Ge, Shaokang Wang

机构 * Xidian University(西安电子科技大学) Shanghai Jiaotong University(上海交通大学)

专题命中 文生图 :text-to-image(abstract);diffusion(abstract);inpainting(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06220 2025-10-31 cs.CV cs.AI 77%

GenIR: Generative Visual Feedback for Mental Image Retrieval

Diji Yang, Minghao Liu, Chung-Hsiang Lo, Yi Zhang, James Davis

机构 * University of California Santa Cruz(加州大学圣克ruz分校) Northeastern University(东北大学) Accenture(Accenture公司)

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02631 2025-10-31 cs.HC cs.AI 50%

Reflection on Data Storytelling Tools in the Generative AI Era from the Human-AI Collaboration Perspective

Haotian Li, Yun Wang, Huamin Qu

机构 * Microsoft Research Asia(微软亚洲研究院) The Hong Kong University of Science and Technology(香港科技大学)

专题命中 文生图 :text-to-image(abstract)

Comments This paper is a sequel to the CHI 24 paper "Where Are We So Far? Understanding Data Storytelling Tools from the Perspective of Human-AI Collaboration (https://doi.org/10.1145/3613904.3642726), aiming to refresh our understanding with the latest advancements. It is accepted at IEEE VIS 25

详情

展开后加载摘要…

URL PDF HTML 收藏