机构
*
Dept. of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(香港理工大学电子与电气工程系)
;
Speech, Language, and Cognition Laboratory, The University of Hong Kong(香港大学语音、语言与认知实验室)
Large Language Models as Unified Multimodal Learners for Clinical Prediction
作为临床预测统一多模态学习者的大语言模型
Ajay Madhavan Ravichandran, Bilgin Osmandoja, Klemens Budde, Klaus Netter, Tobias Strapatsas, Aljoscha Burchardt, Sebastian Möller, Roland Roller
机构
*
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心(DFKI))
;
Charité Universitätsmedizin Berlin(柏林夏里特大学医学中心)
;
DNC Information Management GmbH(DNC信息管理有限公司)
;
Klinik für Akut- und Notfallmedizin, Asklepios Klinikum Harburg(哈堡阿斯克勒庇俄斯医院急性与急诊医学科)
;
Technical University Berlin(柏林工业大学)
OmniFocus: Query-Guided Modality-Balanced Token Compression for Omni-Modal Large Language Models
OmniFocus:用于全模态大语言模型的查询引导模态平衡令牌压缩
Shijie Cao, Qingyu Zhang, Boxi Yu, Yuzhong Zhang, Boxi Cao, Yaojie Lu, Hongyu Lin, Xianpei Han, Le Sun
机构
*
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉科学学院)
;
Chinese Information Processing Laboratory, Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所中文信息处理实验室)
;
University of Limerick(利默里克大学)
;
CUHK, Shenzhen(香港中文大学(深圳))
Cross4D-JEPA: Dense Cross-modal Correspondence Distillation for 4D Point Cloud Representation Learning
Cross4D-JEPA: 用于4D点云表示学习的密集跨模态对应蒸馏
Trung Thanh Nguyen, Hai Nguyen-Truong, Tu Vo, Hoang M. Truong, Tuan-Anh Vu
机构
*
Nagoya University(名古屋大学)
;
Northeastern University(东北大学)
;
KC Machine Learning Lab(KC机器学习实验室)
;
University of Science, Vietnam National University Ho Chi Minh City(胡志明市越南国立大学理科大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴集团达摩院)
;
Hupan Lab(华平实验室)
;
Huazhong University of Science and Technology(华中科技大学)
;
East China Normal University(华东师范大学)
;
Shanghai Jiao Tong University(上海交通大学)
机构
*
Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)
;
Institute for Embodied Intelligence and Robotics, Tsinghua University(清华大学智能感知与机器人研究院)
;
Department of Computer Science and Engineering, Shanghai Jiao Tong University(上海交通大学计算机科学与工程系)
;
Huakong AI Plus Company Limited(华冠AIplus有限公司)
;
Didi International Business Group(滴滴国际商务集团)
An approach with Visual and Tabular Mamba to multimodal medical data using Mixed Fusion
一种基于视觉和表格Mamba的混合融合多模态医疗数据方法
Matheus B. Rocha, Gustavo B. Dettogni, Renato A. Krohling
机构
*
Labcin - Nature Inspired Computing Lab, Federal University of Esp\'irito Santo, Vit\'oria, Brazil PPGI - Graduate Program in Computer Science, Federal University of Esp\'irito Santo, Vit\'oria, Brazil
Plug-and-Adapt: Multimodal Coreference Resolution at First Sight with a Pretrained Alignment Model
即插即适应:基于预训练对齐模型的首眼多模态指代消解
Jinghan Wu, Jing Li, Ivor W. Tsang, Xuetao Zhang
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi'an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能全国重点实验室)
;
Centre for Frontier AI Research and Institute of High-Performance Computing, Agency for Science, Technology and Research (A*STAR)(新加坡科技研究局前沿人工智能研究中心与高性能计算研究所)
Attention, not scale, drives human-AI alignment in multimodal language prediction
注意力,而非规模,驱动多模态语言预测中的人机对齐
Viktor Kewenig, Andrew Lampinen, Samuel A. Nastase, Christopher Edwards, Quitterie Lacome D'Elascombe, Akilles Rechardt, Jeremy I Skipper, Gabriella Vigliocco
机构
*
Psychology and Language Science, Experimental Psychology, University College London, London, UK(心理学与语言科学、实验心理学,伦敦大学学院,伦敦,英国)
;
Google Deepmind, Mountain View, US(谷歌DeepMind,山景城,美国)
;
Princeton Neuroscience Institute, Princeton University, Princeton, NJ, USA(普林斯顿神经科学研究所,普林斯顿大学,普林斯顿,新泽西州,美国)
;
Computer Science Department, Exeter University(计算机科学系,埃克塞特大学)
Information-Theoretic Decomposition for Multimodal Interaction Learning
多模态交互学习的信息论分解
Zequn Yang, Yake Wei, Haotian Ni, Zhihao Xu, Di Hu
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China, Beijing(中国人民大学人工智能学院,北京)
;
Beijing Key Laboratory of Research on Large Models(北京大模型研究关键实验室)
;
Engineering Research Center of Next-Generation Intelligent Search(下一代智能搜索与推荐工程研究中心)
;
Beihang University, Beijing(北航,北京)
;
Gaotu Techedu Inc.(高图科技有限公司)
机构
*
Department of Computer Science and Technology, Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)计算机科学与技术学院)
;
School of Intelligence Science and Engineering, College of Artificial Intelligence, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)人工智能学院智能科学与工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Guangdong Key Laboratory of Biomedical Measurements and Ultrasound Imaging, School of Biomedical Engineering, Shenzhen University Medical School, Shenzhen University(深圳大学医学部生物医学工程学院广东省生物医学测量与超声成像重点实验室)
;
Department of Radiology, The People’s Hospital of Guangxi Zhuang Autonomous Region, Guangxi Academy of Medical Sciences(广西壮族自治区人民医院放射科,广西医学科学院)
;
Shenzhen Sixth People’s Hospital (Nanshan Hospital), Huazhong University of Science and Technology Union Shenzhen Hospital(华中科技大学协和深圳医院(深圳市第六人民医院))
;
School of Basic Medical Sciences, Shenzhen University(深圳大学基础医学院)
;
Egypt-Japan University of Science and Technology (E-JUST)(埃及日本科技大学)
;
School of Biomedical Engineering, National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, Guangdong Key Laboratory for Biomedical Measurements and Ultrasound Imaging, Shenzhen University Medical School(深圳大学医学部生物医学工程学院,国家地方联合医学超声关键技术工程实验室,广东省生物医学测量与超声成像重点实验室)
机构
*
Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学)
;
Pengcheng Laboratory(鹏城实验室)
;
Ant Group(蚂蚁集团)
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室(深圳))
;
University of Pennsylvania(宾夕法尼亚大学)