Multimodal Fusion with LLMs for Engagement Prediction in Natural Conversation
Cheng Charles Ma, Kevin Hyekang Joo, Alexandria K. Vail, Sunreeta Bhattacharya, Álvaro Fernández García, Kailana Baker-Matsuoka, Sheryl Mathew, Lori L. Holt, Fernando De la Torre
机构
*
Computer Science Department, Carnegie Mellon University(卡内基梅隆大学计算机科学系)
;
Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所)
;
Neuroscience Institute, Carnegie Mellon University(卡内基梅隆大学神经科学研究所)
;
Department of Psychology, The University of Texas at Austin(德克萨斯大学奥斯汀分校心理学系)
;
Center for Perceptual Systems, The University of Texas at Austin(德克萨斯大学奥斯汀分校感知系统中心)
Cross-Enhanced Multimodal Fusion of Eye-Tracking and Facial Features for Alzheimer's Disease Diagnosis
Yujie Nie, Jianzhang Ni, Yonglong Ye, Yuan-Ting Zhang, Yun Kwok Wing, Xiangqing Xu, Xin Ma, Lizhou Fan
机构
*
School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学)
;
Engineering Research Center of Intelligent Unmanned System, Ministry of Education(智能无人机系统工程研究中心,教育部)
;
Department of Psychiatry, The Chinese University of Hong Kong(心理学系,香港中文大学)
;
Department of Electronic Engineering, The Chinese University of Hong Kong(电子工程系,香港中文大学)
;
AICARE Lab, Guangdong Medical University(AICARE实验室,广东医科大学)
;
Department of Neurology, Shandong University of Traditional Chinese Medicine Affiliated Hospital(神经内科,山东中医药大学附属医院)
L2RSI: Cross-view LiDAR-based Place Recognition for Large-scale Urban Scenes via Remote Sensing Imagery
Ziwei Shi, Xiaoran Zhang, Wenjing Xu, Yan Xia, Yu Zang, Siqi Shen, Cheng Wang
机构
*
Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University, China(福建智能城市感知与计算重点实验室,厦门大学,中国)
;
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University, China(多媒体可信感知与高效计算重点实验室,中华人民共和国教育部,厦门大学,中国)
;
University of Science and Technology of China, China(中国科学技术大学,中国)