Leveraging AI multimodal geospatial foundation models for improved near-real-time flood mapping at a global scale
利用AI多模态地理空间基础模型实现全球范围内的改进型实时洪水制图
Mirela G. Tulbure, Julio Caineta, Mark Broich, Mollie D. Gaines, Philippe Rufin, Leon-Friedrich Thomas, Hamed Alemohammad, Jan Hemmerling, Patrick Hostert
CommentsThis manuscript is withdrawn to allow for substantial expansion and restructuring. Based on recent research progress, we plan to add Generalization experiment and reorganize the manuscript structure to improve readability and logical flow. Thank you for your understanding and support
Leveraging Biomolecule and Natural Language through Multi-Modal Learning: A Survey
利用生物分子和自然语言通过多模态学习:一篇综述
Qizhi Pei, Zhimeng Zhou, Kaiyuan Gao, Jinhua Zhu, Yue Wang, Zun Wang, Tao Qin, Lijun Wu, Rui Yan
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)
;
Zhejiang University(浙江大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Zhongguancun Academy(中关村学院)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
利用多模态语言模型评估婴儿的视觉和语言经验一致性
Alvin Wei Ming Tan, Jane Yang, Tarun Sepuri, Khai Loong Aw, Robert Z. Sparks, Zi Yin, Virginia A. Marchman, Michael C. Frank, Bria Long
机构
*
Department of Psychology, Stanford University(心理学系,斯坦福大学)
;
Department of Psychology, University of California, San Diego(心理学系,加州大学圣地亚哥分校)
;
Department of Psychology, Tsinghua University(心理学系,清华大学)
AdaTok: Adaptive Token Compression with Object-Aware Representations for Efficient Multimodal LLMs
AdaTok: 一种基于对象感知表示的自适应令牌压缩方法,用于高效多模态大语言模型
Xinliang Zhang, Lei Zhu, Hangzhou He, Shuang Zeng, Ourui Fu, Jiakui Hu, Zhengjian Yao, Yanye Lu
机构
*
Institute of Medical Technology, Peking University Health Science Center(北京大学医学部医学技术研究所)
;
Department of Biomedical Engineering, Peking University(北京大学生物医学工程系)
;
National Biomedical Imaging Center, Peking University(北京大学国家生物医学成像中心)
Multimodal Fusion with LLMs for Engagement Prediction in Natural Conversation
Cheng Charles Ma, Kevin Hyekang Joo, Alexandria K. Vail, Sunreeta Bhattacharya, Álvaro Fernández García, Kailana Baker-Matsuoka, Sheryl Mathew, Lori L. Holt, Fernando De la Torre
机构
*
Computer Science Department, Carnegie Mellon University(卡内基梅隆大学计算机科学系)
;
Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所)
;
Neuroscience Institute, Carnegie Mellon University(卡内基梅隆大学神经科学研究所)
;
Department of Psychology, The University of Texas at Austin(德克萨斯大学奥斯汀分校心理学系)
;
Center for Perceptual Systems, The University of Texas at Austin(德克萨斯大学奥斯汀分校感知系统中心)
Cross-Enhanced Multimodal Fusion of Eye-Tracking and Facial Features for Alzheimer's Disease Diagnosis
Yujie Nie, Jianzhang Ni, Yonglong Ye, Yuan-Ting Zhang, Yun Kwok Wing, Xiangqing Xu, Xin Ma, Lizhou Fan
机构
*
School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学)
;
Engineering Research Center of Intelligent Unmanned System, Ministry of Education(智能无人机系统工程研究中心,教育部)
;
Department of Psychiatry, The Chinese University of Hong Kong(心理学系,香港中文大学)
;
Department of Electronic Engineering, The Chinese University of Hong Kong(电子工程系,香港中文大学)
;
AICARE Lab, Guangdong Medical University(AICARE实验室,广东医科大学)
;
Department of Neurology, Shandong University of Traditional Chinese Medicine Affiliated Hospital(神经内科,山东中医药大学附属医院)
Enhancing Osteoporosis Detection: An Explainable Multi-Modal Learning Framework with Feature Fusion and Variable Clustering
Mehdi Hosseini Chagahi, Saeed Mohammadi Dashtaki, Niloufar Delfan, Nadia Mohammadi, Farshid Rostami Pouria, Behzad Moshiri, Md. Jalil Piran, Oliver Faust
机构
*
School of Electrical and Computer Engineering, College of Engineering, University of Tehran(塔里班大学电气与计算机工程学院)
;
Department of Epidemiology, Shiraz University of Medical Science(谢尔兹医学科学大学流行病学系)
;
Department of Computer Science and Engineering, Sejong University(世宗大学计算机科学与工程系)
;
School of Computing and Information Science, Anglia Ruskin University(安格利亚 Ruskin 大学计算与信息科学学院)
机构
*
Fudan University(复旦大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
The University of Hong Kong(香港大学)
;
Shenzhen University(深圳大学)
;
University of Science and Technology of China(中国科学技术大学)