RetouchLLM: Training-free Code-based Image Retouching with Vision Language Models
机构 * POSTECH ; Huawei London Research Center(华为伦敦研究中心) ; KAIST(韩国科学技术院) ; Imperial College London(伦敦帝国理工学院)
专题命中 其他VLM :vision language model(title);分类 cs.CV
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * POSTECH ; Huawei London Research Center(华为伦敦研究中心) ; KAIST(韩国科学技术院) ; Imperial College London(伦敦帝国理工学院)
专题命中 其他VLM :vision language model(title);分类 cs.CV
机构 * CSSE Department, Auburn University(计算机科学与工程系,阿伯丁大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted to ICCV 2025. 17 pages, 6 figures, 3 tables
机构 * Max Planck Institute for Informatics(马克斯·普朗克研究所信息学研究所) ; EPFL(瑞士联邦理工学院)
专题命中 其他VLM :vision language model(abstract);分类 cs.LG
Comments SIGGRAPH Asia 2024 Courses. arXiv admin note: text overlap with arXiv:2208.11970 by other authors
Journal ref SIGGRAPH Asia 2024 Courses, Article No.: 8, Pages 1 - 27