SERVAL: Surprisingly Effective Zero-Shot Visual Document Retrieval Powered by Large Vision and Language Models
专题命中 文生图 :text-to-image(abstract)
Comments Accepted
Journal ref EMNLP 2025
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :text-to-image(abstract)
Comments Accepted
Journal ref EMNLP 2025
专题命中 文生图 :text-to-image(abstract)
Comments The paper has been withdrawn by the authors because the current experimental results are not sufficiently reliable. Further optimization and refinement of the methodology are required before the work can be disseminated
机构 * Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系) ; Independent Researcher(独立研究者)
专题命中 文生图 :text-to-image(abstract)
Comments accepted to ICML 2025
专题命中 文生图 :image synthesis(abstract)
机构 * University of Vienna(维也纳大学) ; University of Texas at Austin(德克萨斯大学奥斯汀分校) ; University of Maine(缅因大学) ; McGill University(麦吉尔大学) ; University of Wisconsin(威斯康星大学)
专题命中 文生图 :text-to-image(abstract)
机构 * Archimedes,Athena Reaserch Center, Greece(阿基米德、阿泰纳研究中心) ; National Technical University of Athens, Greece(雅典技术大学) ; University of Crete, Greece(克里特大学)
专题命中 文生图 :image synthesis(abstract)
Comments ICML 2025
机构 * Northeastern University(东北大学) ; University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
专题命中 文生图 :text-to-image(abstract)
机构 * TU Darmstadt(图宾根大学) ; DFKI(德意志联邦人工智能研究中心) ; CERTAIN(CERTAIN公司) ; Centre for Cognitive Science, Darmstadt(达姆施塔特认知科学中心)
专题命中 文生图 :text-to-image(abstract)
专题命中 文生图 :image synthesis(abstract)
Comments 30 pages, comments and suggestions are welcome
专题命中 文生图 :text-to-image(abstract)
Comments ACM C&C 2025. Code available at https://github.com/mkremins/fuzzy-linkography
专题命中 文生图 :image synthesis(abstract)
Comments 19 pages, 10+1 figures, accepted by ApJ
机构 * New York University(纽约大学) ; École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院)
专题命中 文生图 :text-to-image(abstract)
机构 * Carnegie Mellon University(卡内基梅隆大学) ; Northwestern University(西北大学) ; University of Washington(华盛顿大学)
专题命中 文生图 :text-to-image(abstract)
Comments Accepted to ACL Findings 2025
机构 * Simon Fraser University(西蒙弗雷泽大学)
专题命中 文生图 :image synthesis(abstract)
机构 * Northeastern University(东北大学) ; Fujitsu Research of America(富士通美国研究)
专题命中 文生图 :text-to-image(abstract)
Journal ref Proceedings of the 17th Conference on Creativity \& Cognition (C\&C), June 23-25, 2025, Virtual, United Kingdom
专题命中 文生图 :text-to-image(abstract)
Comments 28 pages, 9 figures, 2 interactive figures
机构 * Department of Computer Science and Engineering, National Institute of Technology Durgapur, India(印度德瓦格普国家理工学院计算机科学与工程系) ; Department of Computer Science and Engineering, Indian Institute of Technology Bombay, India(印度孟买印度理工学院计算机科学与工程系) ; Electronics and Communication Sciences Unit, Indian Statistical Institute, Kolkata, India(印度统计研究所加尔各答电子与通信科学单元)
专题命中 文生图 :image synthesis(abstract)
Comments 7 pages, 2 figures, 3 tables
专题命中 文生图 :text-to-image(abstract)
Journal ref Proceedings of the 40th ACM/SIGAPP Symposium on Applied Computing (SAC'25), March 31--April 4, 2025, Catania, Italy
专题命中 文生图 :image synthesis(abstract)
专题命中 文生图 :text-to-image(abstract)
Comments accepted to CHI LBW 2025
专题命中 文生图 :text-to-image(abstract)
Comments 8 pages, 7 figures, 2 tables
专题命中 文生图 :text-to-image(abstract)
专题命中 文生图 :image synthesis(abstract)
Comments 6 pages, 8 figures, published in MNRAS Letters
Journal ref Monthly Notices of the Royal Astronomical Society Letters (2025), 538, 1, L62-L68
专题命中 文生图 :text-to-image(abstract)
Comments 20 pages, 5 figures
专题命中 文生图 :text-to-image(abstract)
专题命中 文生图 :image synthesis(abstract)
Comments NEJLT accpeted
专题命中 文生图 :text-to-image(abstract)
专题命中 文生图 :image synthesis(abstract)
Journal ref 2024 IEEE International Conference on Acoustics, Speech and Signal Processing
专题命中 文生图 :image synthesis(abstract)
专题命中 文生图 :text-to-image(abstract)
Comments 20 pages, 7 figures