Taming Text-to-Image Synthesis for Novices: User-centric Prompt Generation via Multi-turn Guidance
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract)
Comments Accepted by EMNLP 2025 main
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :text-to-image(title,abstract);image synthesis(title,abstract)
Comments Accepted by EMNLP 2025 main
机构 * Institute for AI Industry Research, Tsinghua University(人工智能产业研究院,清华大学) ; School of Software Engineering, Nankai University(软件工程学院,南开大学) ; Department of Computing, Hong Kong Polytechnic University(计算机学院,香港理工大学) ; AsiaInfo Technologies(亚信息科技) ; China-Austria Belt and Road Joint Laboratory on Artificial Intelligence and Advanced Manufacturing, Hangzhou Dianzi University(人工智能与先进制造联合实验室,杭州电子科技大学)
专题命中 文生图 :text-to-image(title,abstract)
机构 * INSERM, LTSI - UMR 1099 University of Rennes(法国里昂大学INSERM LTSI - UMR 1099)
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments https://github.com/vboussot/KonfAI
机构 * School of Information Engineering, Guangdong University of Technology(广东技术大学信息工程学院) ; TikTok, ByteDance Inc(字节跳动) ; School of Computer Science, Anhui University(安徽大学计算机科学学院) ; School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments For the first time, angle-based perception was introduced into the multi-modality image fusion task
专题命中 文生图 :text-to-image(abstract);分类 cs.CV
Comments NeurIPS 2025; 27 pages, 6 figures