RSVLM-QA: A Benchmark Dataset for Remote Sensing Vision Language Model-based Question Answering
机构 * School of Computer Science, University of Technology Sydney(技术悉尼大学计算机科学学院) ; SEDE, University of Technology Sydney(技术悉尼大学SEDE) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
专题命中 视觉问答 :vision language model(title,abstract);VLM(abstract);visual question answering(abstract);分类 cs.CV
Comments This paper has been accepted to the proceedings of the 33rd ACM International Multimedia Conference (ACM Multimedia 2025)