Correlating instruction-tuning (in multimodal models) with vision-language processing (in the brain)
机构 * Technische Universität Berlin(柏林技术大学) ; IIIT Hyderabad(海得拉巴国家理工学院) ; Univ of Maryland(马里兰大学) ; Rice Univ(Rice 大学) ; Univ of Wisconsin - Madison(威斯康星大学麦迪逊分校) ; Spector Inc(Spector 公司) ; Microsoft(微软公司)
专题命中 图文多模态 :multimodal(title,abstract);MLLM(abstract);分类 cs.AI
Comments 30 pages, 22 figures, The Thirteenth International Conference on Learning Representations, ICLR-2025, Singapore. https://openreview.net/pdf?id=xkgfLXZ4e0
Journal ref ICLR-2025, Singapore