Architecting Clinical Collaboration: Multi-Agent Reasoning Systems for Multimodal Medical VQA
机构 * Georgia Institute of Technology Atlanta, GA, USA(佐治亚理工学院)
专题命中 多模态Agent :multimodal(title);分类 cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Georgia Institute of Technology Atlanta, GA, USA(佐治亚理工学院)
专题命中 多模态Agent :multimodal(title);分类 cs.AI
机构 * SnT, University of Luxembourg(斯诺特研究所,卢森堡大学) ; Artec3D, Luxembourg(Artec3D,卢森堡)
专题命中 多模态Agent :multimodal(abstract);分类 cs.CV、cs.AI
机构 * SJTU(上海交通大学) ; EIT(欧洲研究所) ; THU(清华大学) ; Galbot ; PKU(北京大学) ; UIUC(伊利诺伊大学香槟分校) ; USTC(中国科学技术大学)
专题命中 多模态Agent :multimodal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments 9 pages,7 figures,conference
机构 * Brown University(布朗大学) ; University of Warwick(沃里克大学) ; Samsung US(三星美国分公司) ; Boston University(波士顿大学) ; University of Alabama at Birmingham(阿拉巴马大学伯明翰分校) ; Rutgers University(罗格斯大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments COLM 2025, 30 pages, 10 figures, 16 tables
机构 * Department of Computer Science, Virginia Tech, Blacksburg, VA, USA(计算机科学系,弗吉尼亚理工学院,布莱克斯堡,VA,美国) ; Department of Agricultural and Applied Economics, Virginia Tech, Blacksburg, VA, USA(农业与应用经济学系,弗吉尼亚理工学院,布莱克斯堡,VA,美国)
专题命中 多模态训练与对齐 :multimodal(title,abstract)
Journal ref The 2025 Conference on Empirical Methods in Natural Language Processing
机构 * National Engineering Laboratory for Integrated Aero-Space-Ground-Ocean Big Data Application Technology(集成空天地海大数据应用技术国家工程实验室) ; Northwestern Polytechnical University(西北工业大学) ; Huiying Medical Technology Company Ltd.(慧影医疗技术有限公司) ; The School of Computer Science(计算机学院) ; The University of Sydney(悉尼大学) ; Department of Computer Science and Engineering(计算机科学与工程系) ; The Chinese University of Hong Kong(香港中文大学)
专题命中 多模态训练与对齐 :multimodal(abstract);cross-modal(abstract);分类 cs.CV
机构 * University of Southern California(南加州大学) ; Information Sciences Institute, University of Southern California(信息科学研究所,南加州大学)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
Comments To appear at EMNLP 2025 (Findings)
机构 * School of Mathematical Sciences, Tongji University(数学科学学院,同济大学)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV、cs.AI
Comments 8 pages, 3 figures
机构 * Intelligent Creation Lab, ByteDance(字节跳动智能创作实验室)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Homepage: https://omnihuman-lab.github.io/v1_5/
机构 * Beijing Institute of Technology(北京理工大学) ; Zhejiang University(浙江大学)
专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV
Comments ACMMM 2025
机构 * Nanyang Technological University(南洋理工大学) ; Westlake University(西湖大学)
专题命中 多模态训练与对齐 :image-text(abstract);分类 cs.CV
Comments Accepted by CVPR 2025
机构 * 1 Ben Gurion University of the Negev, Software
专题命中 其他多模态 :multimodal(title,abstract);分类 cs.AI
专题命中 其他多模态 :multimodal(abstract);分类 cs.AI
Comments 17 pages, 4 figures, 2 tables