Benchmarking Multimodal LLMs on Recognition and Understanding over Chemical Tables
在化学表格上评估多模态大语言模型的基准测试
机构 * State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室) ; University of Science and Technology of China(中国科学技术大学) ; Artificial Intelligence Research Institute(人工智能研究院) ; iFLYTEK Co., Ltd(iFLYTEK公司)
专题命中 视觉问答 :multimodal large language model(abstract);分类 cs.AI
AI总结 本文提出ChemTable基准,用于评估多模态模型在理解化学表格中的能力,揭示了现有模型在跨模态对齐和领域推理方面的不足。