arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2025-10-14 至 2025-10-14 共收录 7 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 7 篇

2510.10633 2025-10-14 cs.AI 88%

Collaborative Text-to-Image Generation via Multi-Agent Reinforcement Learning and Semantic Fusion

Jiabao Shi, Minfeng Qi, Lefeng Zhang, Di Wang, Yingjie Zhao, Ziying Li, Yalong Xing, Ningran Li

机构 * Minzu University of China(民族大学) City University of Macau(澳门城市大学) Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center (National Supercomputer Center in Jinan), Qilu University of Technology (Shandong Academy of Sciences)(计算能力网络与信息安全重点实验室,教育部,山东计算机科学中心(济南国家超级计算机中心),齐鲁工业大学(山东省科学院)) The University of Adelaide(阿德莱德大学)

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract)

Comments 16 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.08114 2025-10-14 cs.CV 84%

RATLIP: Generative Adversarial CLIP Text-to-Image Synthesis Based on Recurrent Affine Transformations

Chengde Lin, Xijun Lu, Guangxi Chen

机构 * School of Artificial Intelligence, Guangxi Colleges and Universities Key Laboratory of AI Algorithm Engineering(人工智能学院、广西 Colleges and Universities AI 算法工程重点实验室)

专题命中 文生图 :text-to-image(title);image synthesis(title);分类 cs.CV

Comments Accepted by 2024 IEEE International Conference on Systems, Man, and Cybernetics(SMC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01720 2025-10-14 cs.CV cs.GR cs.LG 81%

Generating Multi-Image Synthetic Data for Text-to-Image Customization

Nupur Kumari, Xi Yin, Jun-Yan Zhu, Ishan Misra, Samaneh Azadi

机构 * Carnegie Mellon University(卡内基梅隆大学) Meta

专题命中 文生图 :text-to-image(title,abstract);分类 cs.CV、cs.GR

Comments ICCV 2025. Project webpage: https://www.cs.cmu.edu/~syncd-project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10715 2025-10-14 cs.GR cs.CV 79%

VLM-Guided Adaptive Negative Prompting for Creative Generation

Shelly Golan, Yotam Nitzan, Zongze Wu, Or Patashnik

机构 * Adobe Research(Adobe研究院) Tel Aviv University(特拉维夫大学)

专题命中 文生图 :image generation(abstract);text-to-image(abstract);diffusion(abstract);分类 cs.CV、cs.GR

Comments Project page at: https://shelley-golan.github.io/VLM-Guided-Creative-Generation/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11314 2025-10-14 cs.CL 78%

Template-Based Text-to-Image Alignment for Language Accessibility: A Study on Visualizing Text Simplifications

Belkiss Souayed, Sarah Ebling, Yingqiang Gao

机构 * University of Zurich(苏黎世大学)

专题命中 文生图 :text-to-image(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.06701 2025-10-14 cs.CV cs.AI cs.LG 74%

Camouflaged Image Synthesis Is All You Need to Boost Camouflaged Detection

Haichao Zhang, Can Qin, Yu Yin, Yun Fu

机构 * Department of Electrical and Computer Engineering, Northeastern University(电气与计算机工程系,东北大学) Department of Electrical Engineering and Computer Science, Case Western Reserve University(电气工程与计算机科学系,凯斯西储大学)

专题命中 文生图 :image synthesis(title);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05970 2025-10-14 cs.CV 57%

Automatic Synthesis of High-Quality Triplet Data for Composed Image Retrieval

Haiwen Li, Delong Liu, Zhaohui Hou, Zhicheng Zhao, Fei Su

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 文生图 :text-to-image(abstract);分类 cs.CV

Comments This paper was originally submitted to ACM MM 2025 on April 12, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏