MM-NeuroOnco: A Multimodal Benchmark and Instruction Dataset for MRI-Based Brain Tumor Diagnosis
MM-NeuroOnco: 一种多模态基准和指令数据集用于基于MRI的脑肿瘤诊断
Feng Guo, Jiaxiang Liu, Yang Li, Qianqian Shi, Mingkun Xu
机构
*
Guangdong Institute of Intelligence Science and Technology(广东智能科学与技术研究院)
;
Center for Brain-Inspired Computing Research (CBICR), Department of Precision Instrument, Tsinghua University(脑启发计算研究中心(CBICR)、精密仪器系,清华大学)
Architecture-Agnostic Curriculum Learning for Document Understanding: Empirical Evidence from Text-Only and Multimodal
文档理解的架构无关课程学习:来自纯文本和多模态的实证证据
Mohammed Hamdan, Vincenzo Dentamaro, Giuseppe Pirlo, Mohamed Cheriet
机构
*
1 Synchromedia Laboratory, \' E cole de Technologie Sup\' e rieure (\' E TS), 1100 Notre-Dame St W, Montreal, QC H3C 1K3, Canada
;
2 Department of Computer Science, University of Bari Aldo Moro, Via Orabona 4, 70125 Bari, Italy
机构
*
Shenzhen Technology University(深圳技术大学)
;
Shenzhen University of Information Technology(深圳信息科技学院)
;
University of Washington(华盛顿大学)
;
Foundation Model Team, Meituan(美团基础模型团队)
;
National University of Singapore(新加坡国立大学)
;
Tongji University(同济大学)
FOCA: Frequency-Oriented Cross-Domain Forgery Detection, Localization and Explanation via Multi-Modal Large Language Model
FOCA:基于多模态大语言模型的频率导向跨域伪造检测、定位与解释
Zhou Liu, Tonghua Su, Hongshi Zhang, Fuxiang Yang, Donglin Di, Yang Song, Lei Fan
机构
*
Harbin Institute of Technology(哈尔滨工业大学)
;
DZ-Matrix
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy(广东省人工智能与数字经济实验室)
;
Chongqing Research Institute of HIT(哈尔滨工业大学重庆研究院)
;
University of New South Wales(新南威尔士大学)
CT-Bench: A Benchmark for Multimodal Lesion Understanding in Computed Tomography
CT-Bench:一种用于CT中多模态病变理解的基准测试
Qingqing Zhu, Qiao Jin, Tejas S. Mathai, Yin Fang, Zhizheng Wang, Yifan Yang, Maame Sarfo-Gyamfi, Benjamin Hou, Ran Gu, Praveen T. S. Balamuralikrishna, Kenneth C. Wang, Ronald M. Summers, Zhiyong Lu
FMMD: A multimodal open peer review dataset based on F1000Research
FMMD:基于F1000Research的多模态开放同行评审数据集
Zhenzhen Zhuang, Yuqing Fu, Jing Zhu, Zhangping Zhou, Jialiang Lin
机构
*
School of Computer Science and Engineering, Guangzhou Institute of Science and Technology(计算机科学与工程学院,广州科学与技术研究所)
;
College of Foreign Languages and Cultures, Xiamen University(外语学院,厦门大学)
;
Science and Education Evaluation Lab, Guangzhou Institute of Science and Technology(科学与教育评估实验室,广州科学与技术研究所)
RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
RAVENEA:多模态检索增强视觉文化理解的基准
Jiaang Li, Yifei Yuan, Wenyan Li, Mohammad Aliannejadi, Daniel Hershcovich, Anders Søgaard, Ivan Vulić, Wenxuan Zhang, Paul Pu Liang, Yang Deng, Serge Belongie
机构
*
University of Copenhagen(哥本哈根大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Amsterdam(阿姆斯特丹大学)
;
University of Cambridge(剑桥大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Singapore University of Technology and Design(新加坡科技设计大学)
;
Singapore Management University(新加坡管理大学)
Augmented Intelligence for Multimodal Virtual Biopsy in Breast Cancer Using Generative Artificial Intelligence
增强智能用于乳腺癌多模态虚拟活检的生成式人工智能
Aurora Rofena, Claudia Lucia Piccolo, Bruno Beomonte Zobel, Paolo Soda, Valerio Guarrasi
机构
*
Department of Radiology, Fondazione Policlinico Campus Bio-Medico(放射学系,政策临床医学院)
;
Department of Radiology, Università Campus Bio-Medico di Roma(放射学系,罗马生物医学大学)
;
Department of Diagnostics and Intervention, Radiation Physics, Biomedical Engineering, Umeå University(诊断与介入系,辐射物理,生物医学工程,乌梅拉大学)
Visual Reasoning Benchmark: Evaluating Multimodal LLMs on Classroom-Authentic Visual Problems from Primary Education
视觉推理基准:评估多模态大语言模型在小学课堂真实视觉问题上的能力
Mohamed Huti, Alasdair Mackintosh, Amy Waldock, Dominic Andrews, Maxime Lelièvre, Moritz Boos, Tobias Murray, Paul Atherton, Robin A. A. Ince, Oliver G. B. Garrod
机构
*
College of Computer Science, Chongqing University(重庆大学计算机科学学院)
;
Heavy Rainfall Research Center of China(中国强降水研究中心)
;
CMA Key Open Laboratory of Transforming Climate Resources to Economy, Chongqing Institute of Meteorological Sciences(中国气象局气候资源转化经济重点开放实验室、重庆气象科学研究院)
;
Chongqing Metropolitan College of Science(重庆科学理工学院)
;
School of Intelligence Science(智能科学学院)
;
College of Artificial Intelligence, Harbin Institute of Technology(人工智能学院、哈尔滨工业大学)
;
BNU-UIC Institute of Artificial Intelligence(北京师范大学-国际学院人工智能与未来网络研究院)
;
Future Networks, Beijing Normal University at Zhuhai(未来网络、北京师范大学珠海分校)
A Unified Multimodal Framework for Dataset Construction and Model-Based Diagnosis of Ameloblastoma
一种统一的多模态框架用于数据集构建和基于模型的ameloblastoma诊断
Ajo Babu George, Anna Mariam John, Athul Anoop, Balu Bhasuran
机构
*
DiceMed
;
School of Sciences (SOS), Indira Gandhi National Open University(印度甘地国家开放大学科学学院)
;
School of Information, Florida State University(佛罗里达州立大学信息学院)
Med-MMFL: A Multimodal Federated Learning Benchmark in Healthcare
Med-MMFL:医疗领域的多模态联邦学习基准
Aavash Chhetri, Bibek Niroula, Pratik Shrestha, Yash Raj Shrestha, Lesley A Anderson, Prashnna K Gyawali, Loris Bazzani, Binod Bhattarai
机构
*
University of Aberdeen(阿伯丁大学)
;
NepAl Applied Mathematics and Informatics Institute for research, Nepal(尼泊尔应用数学与信息学研究所)
;
University of Lausanne(洛桑大学)
;
West Virginia University(西弗吉尼亚大学)
;
University of Verona(威尼斯大学)
;
University College London(伦敦大学学院)