An Empirical Study of Federated Prompt Learning for Vision Language Model
机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) ; Zhongguancun Laboratory, China(中关村实验室)
专题命中 幻觉与鲁棒性 :vision language model(title,abstract);VLM(abstract);分类 cs.LG
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) ; Zhongguancun Laboratory, China(中关村实验室)
专题命中 幻觉与鲁棒性 :vision language model(title,abstract);VLM(abstract);分类 cs.LG
机构 * Department of Computer Vision, Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(计算机视觉系,Mohamed bin Zayed人工智能大学) ; Department of Computer Science and Engineering, Michigan State University (MSU)(计算机科学与工程系,密歇根州立大学)
专题命中 幻觉与鲁棒性 :vision-language model(title,abstract);分类 cs.CV
Comments Accepted in BMVC 2025
机构 * Southern University of Science and Technology(南方科技大学) ; City University of Hong Kong(香港城市大学) ; George Mason University(乔治·马歇尔大学)
专题命中 幻觉与鲁棒性 :vision-language model(title);分类 cs.AI、cs.LG
Comments 11 pages, 7 figures, 1 table, accepted to IEEE VIS 2025 (IEEE Transactions on Visualization and Computer Graphics)
专题命中 幻觉与鲁棒性 :grounding(abstract)