AutoV: Loss-Oriented Ranking for Visual Prompt Retrieval in LVLMs
AutoV:面向视觉提示检索的损失导向排名用于大视觉-语言模型
机构 * School of Computer Science, Peking University(北京大学计算机学院) ; ByteDance Inc.(字节跳动公司) ; Shanghai Jiao Tong University(上海交通大学)
专题命中 VLM训练与架构 :LLaVA(abstract,abstract_cn);vision-language model(abstract);grounding(abstract);分类 cs.CV
AI总结 AutoV通过损失导向的提示检索提升大视觉-语言模型在图像理解等任务中的性能。
Comments Accepted by ECCV 2026