MaterialFigBENCH: benchmark dataset with figures for evaluating college-level materials science problem-solving abilities of multimodal large language models
MaterialFigBENCH:用于评估多模态大语言模型在大学级材料科学问题解决能力的基准数据集
Michiko Yoshitake, Yuta Suzuki, Ryo Igarashi, Yoshitaka Ushiku, Keisuke Nagato
Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
多模态解耦与耦合网络用于抗干扰的3D目标检测
Rui Ding, Zhaonian Kuang, Yuzhe Ji, Meng Yang, Xinhu Zheng, Gang Hua
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学)
;
Intelligent Transportation Thrust of the Systems Hub, The Hong Kong University of Science and Technology (Guangzhou)(系统枢纽智能交通方向,香港科技大学(广州))
;
Multimodal Experiences Research Lab, Dolby Laboratories(多模态体验研究实验室,Dolby实验室)
机构
*
Department of Computer Science and Engineering, The Chinese University of Hong Kong, Hong Kong, China(计算机科学与工程系,香港中文大学,香港,中国)
;
Institute of Medical Intelligence and XR, The Chinese University of Hong Kong, Hong Kong, China(医学智能与XR研究所,香港中文大学,香港,中国)
How Well Do Multimodal Models Reason on ECG Signals?
多模态模型在心电图信号上的推理能力如何?
Maxwell A. Xu, Harish Haresamudram, Catherine W. Liu, Patrick Langer, Jathurshan Pradeepkumar, Wanting Mao, Sunita J. Ferns, Aradhana Verma, Jimeng Sun, Paul Schmiedmayer, Xin Liu, Daniel McDuff, Emily B. Fox, James M. Rehg
机构
*
University of Illinois Urbana Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Rush University(拉什大学)
;
ETH Zurich(苏黎世联邦理工学院)
;
St. Christopher's Hospital for Children(圣克里斯opher儿童医院)
;
Stanford University(斯坦福大学)
;
University of Washington(华盛顿大学)
;
Google Inc(谷歌公司)
Toward Multimodal Industrial Fault Analysis: A Single-Speed Chain Conveyor Dataset with Audio and Vibration Signals
迈向多模态工业故障分析:一个包含音频和振动信号的单速链式输送机数据集
Zhang Chen, Yucong Zhang, Xiaoxiao Miao, Ming Li
机构
*
Digital Innovation Research Center, Duke Kunshan University(杜克昆山大学数字创新研究中心)
;
School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)人工智能学院)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
;
School of Computer Science, Wuhan University(武汉大学计算机学院)
PhysLLM: Harnessing Large Language Models for Cross-Modal Remote Physiological Sensing
PhysLLM:利用大语言模型进行跨模态远程生理传感
Yiping Xie, Bo Zhao, Mingtong Dai, Jian-Ping Zhou, Yue Sun, Tao Tan, Weicheng Xie, Linlin Shen, Zitong Yu
机构
*
Shenzhen University(深圳大学)
;
Great Bay University(大鹏大学)
;
National Engineering Laboratory for Big Data System Computing Technology(大数据系统计算技术国家工程实验室)
;
Dongguan Key Laboratory for Intelligence and Information Technology(东莞智能与信息技术重点实验室)
;
Guangdong Medical University(广东医科大学)
;
Southern Medical University (Dongguan People’s Hospital)(南方医科大学(东莞人民医院))
;
Macao Polytechnic University(澳门理工学院)
机构
*
Dalian Maritime University(大连海事大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Tsinghua University(清华大学)
;
Nanyang Technological University(南洋理工大学)
;
Xi'an Jiaotong University(西安交通大学)
;
Renmin University of China(中国人民大学)
;
Wuhan University(武汉大学)
机构
*
Department of Computer Science and Engineering, United International University(计算机科学与工程系,国际大学)
;
Applied Artificial Intelligence and Intelligent Systems (AAIINS) Laboratory(应用人工智能与智能系统实验室)
;
Department of Data Science and Artificial Intelligence, Monash University(数据科学与人工智能系,墨尔本大学)
;
Faculty of Science and Technology, Charles Darwin University(科学与技术学院,查尔斯达尔文大学)
;
Faculty of Arts and Society, Charles Darwin University(艺术与社会学院,查尔斯达尔文大学)
Doxing via the Lens: Revealing Location-related Privacy Leakage on Multi-modal Large Reasoning Models
通过镜头进行Doxing:揭示多模态大推理模型中的位置相关隐私泄露
Weidi Luo, Tianyu Lu, Qiming Zhang, Xiaogeng Liu, Bin Hu, Yue Zhao, Jieyu Zhao, Song Gao, Patrick McDaniel, Zhen Xiang, Chaowei Xiao
机构
*
University of Georgia(佐治亚大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
John Hopkins University(约翰·霍普金斯大学)
;
University of Southern California(南加州大学)
;
University of Maryland, College Park(马里兰大学学院市分校)
CommentsCamera-ready version. Accepted as a poster at the 14th International Conference on Learning Representations (ICLR 2026). For official ICLR page, see https://iclr.cc/virtual/2026/poster/10006914
SportR: A Benchmark for Multimodal Large Language Model Reasoning in Sports
SportR:多模态大语言模型在体育中的推理基准
Haotian Xia, Haonan Ge, Junbo Zou, Hyun Woo Choi, Xuebin Zhang, Danny Suradja, Botao Rui, Ethan Tran, Wendy Jin, Zhen Ye, Xiyang Lin, Christopher Lai, Shengjie Zhang, Junwen Miao, Shichao Chen, Rhys Tracy, Vicente Ordonez, Weining Shen, Hanjie Chen
机构
*
Department of Computer Science, Rice University(Rice大学计算机科学系)
;
Ken Kennedy Institute, Rice University(Rice大学肯尼迪研究所)
;
Department of Statistics, University of California, Irvine(伊利诺伊大学欧文分校统计系)
;
College of Sciences, Georgia Institute of Technology(佐治亚理工学院科学学院)
;
Department of Applied Mathematics and Statistics, Johns Hopkins University(约翰霍普金斯大学应用数学与统计学系)
;
Department of Computer Science, University of California, Santa Barbara(加州大学圣芭芭拉分校计算机科学系)
机构
*
University of Texas(德克萨斯大学)
;
Dell Children’s Medical Center(德尔儿童医学中心)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
University of Nevada, Reno(内华达大学里诺分校)
AgentVista: Evaluating Multimodal Agents in Ultra-Challenging Realistic Visual Scenarios
AgentVista: 评估在超挑战性现实视觉场景中的多模态代理
Zhaochen Su, Jincheng Gao, Hangyu Guo, Zhenhua Liu, Lueyang Zhang, Xinyu Geng, Shijue Huang, Peng Xia, Guanyu Jiang, Cheng Wang, Yue Zhang, Yi R. Fung, Junxian He
机构
*
Hong Kong University of Science and Technology(香港理工大学)
;
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
Yuxuan Yang, Zhonghao Yan, Yi Zhang, Bo Yun, Muxi Diao, Guowei Zhao, Kongming Liang, Wenbin Li, Zhanyu Ma
机构
*
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(人工智能学院,北京邮电大学)
;
Department of Pathology, National Cancer Center/National Clinical Research Center for Cancer/Cancer Hospital, Chinese Academy of Medical Sciences and Peking Union Medical College(pathology department, 国家癌症中心/国家癌症临床研究中心/癌症医院, 中国医学科学院和北京协和医学院)
Chenggang Rong, Tao Han, Zhiyuan Zhao, Yaowu Fan, Jia Wan, Song Guo, Yuan Yuan, Junyu Gao
机构
*
Northwestern Polytechnical University(北华大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究所)
;
Sun Yat-sen University(中山大学)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))