CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
机构 * Chengdu Institute of Computer Applications, Chinese Academy of Sciences(成都计算机应用研究所,中国科学院) ; University of Chinese Academy of Sciences(中国科学院大学) ; International Research Institute for Artificial Intelligence, Harbin Institute of Technology (Shenzhen)(人工智能国际研究院,哈尔滨工业大学(深圳)) ; Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) ; Southwest Jiaotong University(西南交通大学) ; Tencent(腾讯)
专题命中 幻觉与鲁棒性 :visual language model(abstract);分类 cs.CV
Comments ACMMM2025(oral)