CommentsAccepted to the 27th International Society for Music Information Retrieval Conference (ISMIR 2026), Abu Dhabi, UAE. Project page: https://joonhyungbae.github.io/skypiano/
On-Device Inference versus Wireless Streaming: Energy-Efficient Multi-Modal Deep Learning for Wearable Cardiovascular Patches
面向心血管传感器贴片的端到端多模态微型CNN原型设计
Mustafa Fuad Rifet Ibrahim, Tunc Alkanat, Felix Manthey, Maurice Meijer, Alexander Schlaefer, Peer Stelldinger
机构
*
CTO System Innovation, NXP Semiconductors Germany GmbH(NXP半导体德国系统创新部)
;
Advanced Chip Engineering, NXP Semiconductors(NXP半导体先进芯片工程部)
;
Business Line Secure Connected Edge, NXP Semiconductors(NXP半导体安全连接边缘业务线)
;
Institute of Medical Technology and Intelligent Systems, Hamburg University of Technology(汉堡技术大学医学技术与智能系统研究所)
;
Department of Informatics, Hamburg University of Applied Sciences(汉堡应用科学大学信息学院)
Comments16 pages, 2 figures. Extended version of our 2024 IEEE PerCom paper, with direct on-device energy measurements, a BLE communication benchmark, architecture comparisons, and an extended evaluation. Submitted to Pervasive and Mobile Computing; Measurement-method clarifications and minor editorial corrections; results and conclusions unchanged
Ruiqi Wu, Yuang Yao, Tengfei Ma, Chenran Zhang, Na Su, Tao Zhou, Geng Chen, Wen Fan, Yi Zhou
机构
*
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Department of Ophthalmology, The First Affiliated Hospital of Nanjing Medical University(南京医科大学第一附属医院眼科学系)
;
School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
;
School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院)
Towards Reliable Stain Transfer: An Iterative Data-Model Co-Optimization Framework Based on Multimodal Expert-Guided Assessment
迈向可靠的染色转移:基于多模态专家指导评估的迭代数据-模型协同优化框架
Siyuan Xu, Yan Wang, Haofei Song, Lili Gao, Jiansheng Wang, Qing Zhang, Dan Huang, Boxiang Yun, Hongkai Xiong, Qingli Li
机构
*
East China Normal University(华东师范大学)
;
Ruijin Hospital, Shanghai Jiao Tong University School of Medicine(上海交通大学医学院附属瑞金医院)
;
Hangzhou Hyperspectral Imaging Technology Co., Ltd.(杭州高光谱成像技术有限公司)
;
Fudan University Shanghai Cancer Center(复旦大学附属肿瘤医院)
机构
*
Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(西安交通大学人工智能与机器人研究所)
;
Xingchen AGI Lab, China Telecom Artificial Intelligence Technology (Beijing) Co., Ltd(星辰通用人工智能实验室,中国电信人工智能技术(北京)有限公司)
;
University of Science and Technology Beijing(北京科技大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Shanghai Jiao Tong University(上海交通大学)
机构
*
Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)
;
Institute of Medical Intelligence and XR, The Chinese University of Hong Kong(香港中文大学医学智能与扩展现实研究所)
CommentsDataset can be accessed via zenodo DOI https://doi.org/10.5281/zenodo.18483292 For citation use the primary academic reference: Garbe J et al. Towards predicting sedation depth in endoscopy with large clinically annotated EEG data of continuous Propofol sedation. In: P. Andreevetal (Eds.): AIME2026, LNAI 16749, p.1-6, Springer, 2026. https://doi.org/10.1007/978-3-032-30813-9_58
PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation
PixDLM:一种用于无人机推理分割的双路径多模态语言模型
Shuyan Ke, Yifan Mei, Changli Wu, Yonghan Zheng, Jiayi Ji, Liujuan Cao, Rongrong Ji
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学)
Querying Multimodal Scientific Papers with AI: Practices and Preferences Across Blind, Low-Vision, and Sighted Scientists
用人工智能查询多模态科学论文:盲人、低视力和视力正常科学家的实践与偏好
Arnavi Chheda-Kothary, Lucy Lu Wang, Joseph Chee Chang, Jonathan Bragg
机构
*
Paul G. Allen School of Computer Science \& Engineering University of Washington Seattle WA USA
;
The Information School University of Washington Allen Institute for AI Seattle WA USA
;
Allen Institute for AI Seattle WA USA
;
University of Washington
;
Allen Institute for AI
Learning Sparse Latent Predictive Foundation Model for Multimodal Neuroimaging
学习用于多模态神经影像的稀疏潜在预测基础模型
Haoxu Huang, Long Chen, Jingyun Chen, Jinu Hyun, James Ryan Loftus, Kara Melmed, Daniel Orringer, Jennifer Frontera, Seena Dehkharghani, Arjun Masurkar, Narges Razavian
机构
*
New York University, Center for Data Science(纽约大学数据科学中心)
;
NYU Grossman School of Medicine, Department of Radiology(纽约大学格罗斯曼医学院放射学系)
;
State University of New York at Binghamton, School of Computing(纽约州立大学宾汉姆顿分校计算机学院)
;
NYU Grossman School of Medicine, Department of Neurology(纽约大学格罗斯曼医学院神经病学系)
;
NYU Grossman School of Medicine, Department of Neurosurgery(纽约大学格罗斯曼医学院神经外科学系)
;
NYU Grossman School of Medicine, Department of Pathology(纽约大学格罗斯曼医学院病理学系)
;
School of Medicine, Department of Radiology, Stanford(斯坦福大学医学院放射学系)
;
NYU Grossman School of Medicine, Department of Neuroscience(纽约大学格罗斯曼医学院神经科学系)
;
NYU Grossman School of Medicine, Neuroscience Institute(纽约大学格罗斯曼医学院神经科学研究所)
Multi-Modal Semantic Segmentation of Electrolyzer Components for Sustainable Hydrogen Technologies: A Dual-Branch Deep Learning Approach
用于可持续氢能技术的电解槽组件多模态语义分割:一种双分支深度学习方法
Wasimul Karim, Nur Mohammad Fahad, Abdul Hasib Siddique, Md Rafiqul Islam, Hooman Mehdizadeh-Rad, Asif Karim, Sami Azam
机构
*
Applied Artificial Intelligence and INtelligent Systems (AAIINS) Laboratory(应用人工智能与智能系统(AAIINS)实验室)
;
University of Scholars(学者大学)
;
Murdoch University(莫道克大学)
;
Charles Darwin University(查尔斯达尔文大学)
Paired Uterine Whole-Slide Images and Pathology Reports for Multimodal Computational Pathology
用于多模态计算病理学的配对子宫全切片图像和病理报告
Han Li, Jingsong Liu, Ayako Ura, Junlin Hou, Zhengyang Xu, Azar Kazemi, Oskar Thaeter, Christian Grashei, Fabian Gülhan, Reza Nasirigerdeh, Xun Ma, Rui Yan, Hao Chen, S. Kevin Zhou, Nassir Navab, Carolin Mogler, Peter Schüffler
机构
*
Institute of Pathology, Technical University of Munich(慕尼黑工业大学病理研究所)
;
Computer Aided Medical Procedures (CAMP), Technical University of Munich(慕尼黑工业大学计算机辅助医疗程序(CAMP))
;
Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心)
;
Department of Human Pathology, Juntendo University Graduate School of Medicine(顺天堂大学医学研究生院人体病理学部)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Munich Data Science Institute (MDSI)(慕尼黑数据科学研究所)
TikStance: A Multimodal and Hierarchical Dataset for Multi-target Stance Analysis in TikTok Political Conversations
TikStance:用于TikTok政治对话中多目标立场分析的多模态分层数据集
Yazhi Zhang, Fuqiang Niu, Bowen Zhang
机构
*
School of Artificial Intelligence, Shenzhen Technology University(深圳技术大学人工智能学院)
;
School of Cyber Science and Technology, University of Science and Technology of China(中国科学技术大学网络空间安全学院)