机构
*
University of International Relations(国际关系学院)
;
Kedge Business School(凯致商学院)
;
Peking University(北京大学)
;
Tsinghua University(清华大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Capital Normal University(首都师范大学)
HEad and neCK TumOR (HECKTOR) 2025: Benchmark of Segmentation, Diagnosis, and Prognosis in Multimodal PET/CT
头颈肿瘤 (HECKTOR) 2025 挑战赛:多模态 PET/CT 中的分割、诊断与预后基准
Numan Saeed, Salma Hassan, Shahad Hardan, Lishan Cai, Xinglong Liang, Moona Mazher, Abdul Qayyum, Yansong Bu, Mengye Lyu, Yue Lin, Mingyuan Meng, Chuanyi Huang, Lisheng Wang, Dalal Chamseddine, Shamimeh Ahrari, Beining Wu, Yifei Chen, Fuyou Mao, Hao Zhang, Baixiang Zhao, Surajit Ray, Muzi Guo, Lei Xiang, Jakob Dexl, Michael Ingrisch, Adrien Depeursinge, Arman Rahmim, Mathieu Hatt, Vincent Andrearczyk, Mohammad Yaqub
机构
*
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)
;
Amsterdam UMC(阿姆斯特丹大学医学中心)
;
The Netherlands Cancer Institute(荷兰癌症研究所)
;
Radboud University Medical Centre(拉德堡德大学医学中心)
;
University College London(伦敦大学学院)
;
Imperial College London(帝国理工学院)
;
Shenzhen Technology University(深圳技术大学)
;
Shenzhen University(深圳大学)
;
Newland Digital Technology(新大陆数字技术)
;
The University of Sydney(悉尼大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University Hospital, Nantes(南特大学医院)
;
Nantes Université, Centrale Nantes, CNRS, LS2N(南特大学、南特中央理工学院、法国国家科学研究中心、LS2N实验室)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
Tsinghua University(清华大学)
;
Central South University(中南大学)
;
University of Glasgow(格拉斯哥大学)
;
China Mobile System Integration Co., Ltd.(中移系统集成有限公司)
;
Subtle Medical Inc.(Subtle Medical公司)
;
University Hospital, LMU Munich(慕尼黑大学医院)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
BC Cancer Research Institute(不列颠哥伦比亚癌症研究所)
;
HES-SO Valais-Wallis University of Applied Sciences and Arts(HES-SO瓦莱州应用科学与艺术大学)
;
Lausanne University Hospital (CHUV)(洛桑大学医院)
;
LaTIM, INSERM, UMR 1101, Univ Brest(LaTIM实验室、法国国家健康与医学研究院、UMR 1101、布雷斯特大学)
Comments17 pages, 4 figures, 4 tables. Overview paper for the HECKTOR 2025 challenge, held as a satellite event at MICCAI 2025. Challenge website: https://hecktor.grand-challenge.org/
MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning
MathVis-Fine:通过渐进式依赖引导训练将视觉监督与必要性对齐的多模态数学推理
Wanshi Xu, Haokun Zhao, Haidong Yuan, Songjun Cao, Long Ma
机构
*
School of ECE, Peking University(北京大学电子与计算机工程学院)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与技术学院)
;
School of Software and Microelectronics, Peking University(北京大学软件与微电子学院)
;
Tencent Youtu Lab(腾讯优图实验室)
Comments8 pages, 6 figures. To appear in Proceedings of the 8th International Workshop on IoT Applications and Industry 5.0 (IoTI5 2026), co-located with IEEE DCOSS-IoT 2026, Reykjavik, Iceland, June 2026
Probing, Fusion, and Trustworthiness: A Systematic Evaluation of Foundation Model Representations for Multimodal Cancer Analysis
探测、融合与可信度:基础模型表示在多模态癌症分析中的系统评估
Jingyu Hu, Giuseppe Tripodi, Reed Naidoo, Sarah F. McGough, Tapabrata Chakraborti
机构
*
The Alan Turing Institute(艾伦·图灵研究所)
;
University of Bristol(布里斯托大学)
;
University of Manchester(曼彻斯特大学)
;
The Institute of Cancer Research(癌症研究所)
;
Genentech(基因泰克)
UrbanWell: Benchmarking Multimodal Large Language Models for Spatio-Temporal Urban Wellbeing Analytics
UrbanWell: 面向时空城市福祉分析的多模态大语言模型基准测试
Yanxin Xi, Xiang Su, Jie Feng, Yu Liu, Sasu Tarkoma, Pan Hui
机构
*
University of Helsinki(赫尔辛基大学)
;
Zhongguancun Academy(中关村学院)
;
University of Oxford(牛津大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
NEST3D: A High-Resolution Multimodal Dataset of Sociable Weaver Tree Nests
NEST3D:织布鸟树巢的高分辨率多模态数据集
Constanza A. Molina Catricheo, Simon Boeder, Ting-Jia Guo, Giacomo May, Clément Berthelot, Devis Tuia, Friedrich Fedor Reinhard, Fabio Remondino, Benjamin Risse
机构
*
Institute for Geoinformatics (ifgi), University of Münster(明斯特大学地理信息学研究所)
;
École Polytechnique Fédérale de Lausanne (EPFL)(洛桑联邦理工学院)
;
Max Planck Institute of Animal Behavior(马克斯·普朗克动物行为研究所)
;
University of Konstanz(康斯坦茨大学)
;
Kuzikus Research Station(库兹库斯研究站)
;
Fondazione Bruno Kessler (FBK)(布鲁诺·凯斯勒基金会)
Cross-Modal Benchmarking for Robotic Perception in Natural Environments
自然环境中机器人感知的跨模态基准测试
David Hall, Joshua Knights, Mark Cox, Peyman Moghadam
机构
*
CSIRO Robotics, CSIRO, Australia(CSIRO机器人研究所,CSIRO,澳大利亚)
;
University of Sydney (USyd), Australia(悉尼大学(USyd),澳大利亚)
;
Queensland University of Technology (QUT), Australia(昆士兰理工大学(QUT),澳大利亚)
PereStruct: Multimodal Semantic Assembly for Robust Historical Document Parsing
PereStruct: 面向鲁棒历史文档解析的多模态语义组装
Maksim Shandybo, Ivan Bespalov, Daniil Yefimov, Marina Kosheleva, Alexander Loukianov
机构
*
IGIC RAS(俄罗斯科学院信息传输问题研究所)
;
Yandex Cloud
;
National University of Science and Technology MISIS(莫斯科国立钢铁合金学院)
;
Nekrasov Central Universal Scientific Library(涅克拉索夫中央综合科学图书馆)
From Vision to Text: A Compact Multimodal Approach for Robust, Cross-Domain Presentation Attack Detection on ID Cards
从视觉到文本:一种用于身份证件跨域鲁棒演示攻击检测的紧凑多模态方法
Qingwen Zeng, Juan E. Tapia, Sneha Das, Christoph Busch
机构
*
da/sec-Biometrics and Security Research Group, Hochschule Darmstadt(da/sec生物安全研究组,达姆施塔特应用技术大学)
;
Technical University of Denmark (DTU)(丹麦技术大学(DTU))
Evidence-Based Intelligent Diagnostic and Therapeutic Visualization System with Large Language Models: Multi-Turn Interaction and Multimodal Treatment Plan Generation
基于证据的智能诊断与治疗可视化系统与大语言模型:多轮交互与多模态治疗方案生成
Yunhan Wang, Yuda Wang, Zhiying Tu, Mingqiang Song, Li Song, Kun Li, Dianhui Chu, Bolin Zhang
机构
*
Harbin Institute of Technology, Weihai(哈尔滨工业大学(威海))
;
Harbin Institute of Technology (Weihai) Qingdao Research Institute(哈尔滨工业大学(威海)青岛研究院)
;
Shandong Key Laboratory of Digital Service Computing Technology and Systems(山东省数字服务计算技术与系统重点实验室)
;
Weihai Municipal Hospital(威海市人民医院)
;
Shanghai Taizhu Technology Co., Ltd(上海泰山技术有限公司)
;
Tianjin Zhifu Qihuang Medical Technology Co., Ltd(天津中孚启黄医疗技术有限公司)
Performance of large language models in the optical diagnosis of colorectal polyps
大型语言模型在结直肠息肉光学诊断中的性能
Joshua C. Vences, William T. Tran, Nikko Gimpaya, Catharine M. Walsh, Rishad J. Khan, Robert Bechara, Asher C. Wiggins, Celine N. Rousan, Kaitlyn V. G. L. Morgado, Angie Ibrahim, Kevin H. M. Kuo, Daniel von Renteln, Alexander Hann, Dennis L. Shung, Michael A. Scaffidi, Charles Ménard, Joshua Landy, Samir C. Grover