LISA: A Layer-wise Integration and Suppression Approach for Hallucination Mitigation in Multimodal Large Language Models
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);grounding(abstract);分类 cs.CV
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments [Accepted by AAAI2026] Project Page: https://zju-real.github.io/gui-rcpo Code: https://github.com/zju-real/gui-rcpo
专题命中 视觉定位与Grounding :VLM(abstract);grounding(abstract);分类 cs.CV、cs.AI
机构 * Machine Learning Lab IIIT Hyderabad(IIIT Hyderabad 机器学习实验室) ; Bosch Global Software Technologies(博世全球软件技术公司)
专题命中 视觉定位与Grounding :vision-language model(abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Accepted to AAAI 2026 (Oral), Project Page: https://github.com/JiuTian-VL/SemanticVLA
机构 * City University of Hong Kong(香港城市大学) ; Monash University(墨尔本大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Project Page: https://layerpeeler.github.io/