Don't Learn, Ground: A Case for Natural Language Inference with Visual Grounding
不要学习,而是依托:自然语言推理与视觉依托的案例
机构 * Utrecht University(乌特勒支大学)
专题命中 文生图 :text-to-image(abstract)
AI总结 本文提出了一种基于视觉依托的零样本自然语言推理方法,通过生成视觉表示并比较与假设的相似度,实现高精度推理,展示了对文本偏见的鲁棒性。
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
不要学习,而是依托:自然语言推理与视觉依托的案例
机构 * Utrecht University(乌特勒支大学)
专题命中 文生图 :text-to-image(abstract)
AI总结 本文提出了一种基于视觉依托的零样本自然语言推理方法,通过生成视觉表示并比较与假设的相似度,实现高精度推理,展示了对文本偏见的鲁棒性。
面板式灵魂:一种用于AI辅助漫画创作中表现力面部的表演性工作流程
专题命中 文生图 :text-to-image(abstract)
AI总结 本文提出了一种交互式工作流程,用于AI辅助漫画创作中表现力面部的生成,通过双混合流程实现艺术家意图与AI执行的高效衔接。
Comments NeurIPS 2025 Creative AI Track, The Thirty-Ninth Annual Conference on Neural Information Processing Systems
专题命中 文生图 :text-to-image(abstract)
Comments accepted for publication in the Association for the Advancement of Artificial Intelligence (AAAI), 2026
专题命中 文生图 :text-to-image(abstract)
机构 * Microsoft Research Asia(微软亚洲研究院) ; The Hong Kong University of Science and Technology(香港科技大学)
专题命中 文生图 :text-to-image(abstract)
Comments This paper is a sequel to the CHI 24 paper "Where Are We So Far? Understanding Data Storytelling Tools from the Perspective of Human-AI Collaboration (https://doi.org/10.1145/3613904.3642726), aiming to refresh our understanding with the latest advancements. It is accepted at IEEE VIS 25
机构 * Michigan State University(密歇根州立大学) ; Reality Defender
专题命中 文生图 :text-to-image(abstract)
Comments Accepted as NeurIPS 2025 poster
机构 * Fujitsu Research India(富士通印度研究)
专题命中 文生图 :text-to-image(abstract)
Comments Accepted in EMNLP-2025 Findings
机构 * State Key Laboratory of CAD&CG(计算机辅助设计与图形学国家重点实验室) ; National Key Laboratory for Novel Software Technology(新型软件技术国家实验室)
专题命中 文生图 :text-to-image(abstract)
专题命中 文生图 :text-to-image(abstract)
Comments Accepted
Journal ref EMNLP 2025
专题命中 文生图 :text-to-image(abstract)
Comments The paper has been withdrawn by the authors because the current experimental results are not sufficiently reliable. Further optimization and refinement of the methodology are required before the work can be disseminated
机构 * Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系) ; Independent Researcher(独立研究者)
专题命中 文生图 :text-to-image(abstract)
Comments accepted to ICML 2025
专题命中 文生图 :image synthesis(abstract)
机构 * University of Vienna(维也纳大学) ; University of Texas at Austin(德克萨斯大学奥斯汀分校) ; University of Maine(缅因大学) ; McGill University(麦吉尔大学) ; University of Wisconsin(威斯康星大学)
专题命中 文生图 :text-to-image(abstract)
机构 * Archimedes,Athena Reaserch Center, Greece(阿基米德、阿泰纳研究中心) ; National Technical University of Athens, Greece(雅典技术大学) ; University of Crete, Greece(克里特大学)
专题命中 文生图 :image synthesis(abstract)
Comments ICML 2025
机构 * Northeastern University(东北大学) ; University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
专题命中 文生图 :text-to-image(abstract)
机构 * TU Darmstadt(图宾根大学) ; DFKI(德意志联邦人工智能研究中心) ; CERTAIN(CERTAIN公司) ; Centre for Cognitive Science, Darmstadt(达姆施塔特认知科学中心)
专题命中 文生图 :text-to-image(abstract)
专题命中 文生图 :image synthesis(abstract)
Comments 30 pages, comments and suggestions are welcome
专题命中 文生图 :text-to-image(abstract)
Comments ACM C&C 2025. Code available at https://github.com/mkremins/fuzzy-linkography
专题命中 文生图 :image synthesis(abstract)
Comments 19 pages, 10+1 figures, accepted by ApJ
机构 * New York University(纽约大学) ; École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院)
专题命中 文生图 :text-to-image(abstract)
机构 * Carnegie Mellon University(卡内基梅隆大学) ; Northwestern University(西北大学) ; University of Washington(华盛顿大学)
专题命中 文生图 :text-to-image(abstract)
Comments Accepted to ACL Findings 2025
机构 * Simon Fraser University(西蒙弗雷泽大学)
专题命中 文生图 :image synthesis(abstract)
机构 * Northeastern University(东北大学) ; Fujitsu Research of America(富士通美国研究)
专题命中 文生图 :text-to-image(abstract)
Journal ref Proceedings of the 17th Conference on Creativity \& Cognition (C\&C), June 23-25, 2025, Virtual, United Kingdom
专题命中 文生图 :text-to-image(abstract)
Comments 28 pages, 9 figures, 2 interactive figures
机构 * Department of Computer Science and Engineering, National Institute of Technology Durgapur, India(印度德瓦格普国家理工学院计算机科学与工程系) ; Department of Computer Science and Engineering, Indian Institute of Technology Bombay, India(印度孟买印度理工学院计算机科学与工程系) ; Electronics and Communication Sciences Unit, Indian Statistical Institute, Kolkata, India(印度统计研究所加尔各答电子与通信科学单元)
专题命中 文生图 :image synthesis(abstract)
Comments 7 pages, 2 figures, 3 tables
专题命中 文生图 :text-to-image(abstract)
Journal ref Proceedings of the 40th ACM/SIGAPP Symposium on Applied Computing (SAC'25), March 31--April 4, 2025, Catania, Italy
专题命中 文生图 :image synthesis(abstract)
专题命中 文生图 :text-to-image(abstract)
Comments accepted to CHI LBW 2025
专题命中 文生图 :text-to-image(abstract)
Comments 8 pages, 7 figures, 2 tables
专题命中 文生图 :text-to-image(abstract)