A Multimodal Seq2Seq Transformer for Predicting Brain Responses to Naturalistic Stimuli
Qianyi He, Yuan Chang Leong
机构
*
Data Science Institute University of Chicago(芝加哥大学数据科学研究所)
;
Department of Psychology, Neuroscience Institute University of Chicago(芝加哥大学心理学系、神经科学研究所)
Improving Out-of-distribution Human Activity Recognition via IMU-Video Cross-modal Representation Learning
Seyyed Saeid Cheshmi, Buyao Lyu, Thomas Lisko, Rajesh Rajamani, Robert A. McGovern, Yogatheesan Varatharajah
机构
*
Department of Computer Science & Engineering University of Minnesota(计算机科学与工程系明尼苏达大学)
;
Department of Mechanical Engineering University of Minnesota(机械工程系明尼苏达大学)
;
Department of Neurosurgery University of Minnesota(神经外科系明尼苏达大学)
MCAM: Multimodal Causal Analysis Model for Ego-Vehicle-Level Driving Video Understanding
Tongtong Cheng, Rongzhen Li, Yixin Xiong, Tao Zhang, Jing Wang, Kai Liu
机构
*
Department of Computer Science, Chongqing University, China(重庆大学计算机科学系)
;
National Elite Institute of Engineering, Chongqing University, China(重庆大学工程精英研究院)
;
College of Computer Science and Technology, National University of Deffense Technology, China(国防科技大学计算机科学与技术学院)
What You Have is What You Track: Adaptive and Robust Multimodal Tracking
Yuedong Tan, Jiawei Shao, Eduard Zamfir, Ruanjun Li, Zhaochong An, Chao Ma, Danda Paudel, Luc Van Gool, Radu Timofte, Zongwei Wu
机构
*
TeleAI, China Telecom(TeleAI,中国电信)
;
Computer Vision Lab, CAIDAS & IFI, University of Wurzburg(计算机视觉实验室,CAIDAS与IFI,乌尔姆大学)
;
INSAIT, Sofia University(INSAIT,索菲亚大学)
;
ShanghaiTech University(上海科技大学)
;
University of Copenhagen(哥本哈根大学)
;
AI Institute, Shanghai Jiao Tong University(人工智能研究院,上海交通大学)
机构
*
National Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机科学学院,北京大学)
;
Department of Software and Microelectronics, Peking University(软件与微电子系,北京大学)
;
School of Electronic and Computer Engineering, Peking University(电子与计算机工程学院,北京大学)
;
State Key Laboratory of Virtual Reality Technology and Systems, SCSE, Beihang University(虚拟现实技术与系统国家重点实验室,北航软件学院)
;
Peng Cheng Laboratory(鹏城实验室)
;
Baidu Inc(百度公司)
机构
*
AI Thrust, Information Hub, HKUST(GZ)(香港科技大学(广州)人工智能 thrust 与信息中心)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院)
;
School of Medicine, Stanford University(斯坦福大学医学院)
;
Computer Science Department, UIUC(伊利诺伊大学厄巴纳-香槟分校计算机科学系)
;
Medical Big Data Center, Guangdong Provincial People’s Hospital (Guangdong Academy of Medical Sciences), Southern Medical University(广东省人民医院(广东省医学科学院)医学大数据中心,南方医科大学)
;
Innovation Institute for Artificial Intelligence in Medicine of Zhejiang University, College of Pharmaceutical Sciences, Zhejiang University(浙江大学人工智能医学创新研究院,浙江大学药学院)
;
The Second Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院附属第二医院)
;
GE HealthCare(通用电气医疗)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
IQVIA
;
Computer Science Department, Stanford University(斯坦福大学计算机科学系)
;
Informatics, Harvard Medical School, Harvard University(哈佛医学院,哈佛大学informatics部门)
;
State Key Laboratory for Novel Software Technology at Nanjing University, School of Computer Science, Nanjing University(南京大学新型软件技术国家重点实验室,南京大学计算机科学学院)
机构
*
School of Intelligence Science and Technology, Nanjing University, China(智能科学与技术学院,南京大学)
;
PCA-Lab, School of Computer Science and Engineering, Nanjing University of Science and Technology, China(PCA实验室,计算机科学与工程学院,南京理工大学)
Spatio-Temporal Fuzzy-oriented Multi-Modal Meta-Learning for Fine-grained Emotion Recognition
Jingyao Wang, Wenwen Qiang, Changwen Zheng, Fuchun Sun
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
National Key Laboratory of Space Integrated Information System(空间信息集成系统国家重点实验室)
;
Institute of Software Chinese Academy of Sciences(中国科学院软件研究所)
;
Department of Computer Science and Technology(计算机科学与技术系)
专题命中
视频多模态
:multi-modal(title,abstract);分类 cs.CV
CommentsThis work has been submitted to the IEEE for possible publication
M2WLLM: Multi-Modal Multi-Task Ultra-Short-term Wind Power Prediction Algorithm Based on Large Language Model
Hang Fana, Mingxuan Lib, Zuhan Zhanga, Long Chengc, Yujian Ye, Dunnan Liua
机构
*
School of Economics and Management, North China Electric Power University(华北电力大学经济管理学院)
;
Department of Electrical Engineering, Tsinghua University(清华大学电气工程系)
;
School of Control and Computer Engineering, North China Electric Power University(华北电力大学控制与计算机工程学院)
;
School of Electrical Engineering, Southeast University(东南大学电气工程学院)
机构
*
The University of Tokyo, Graduate School of Arts and Sciences(东京大学艺术与科学研究生院)
;
Center for Information and Neural Networks (CiNet), National Institute of Information and Communications Technology(信息与神经网络中心(CiNet),信息与通信技术国家研究所)
;
The University of Osaka, Graduate School of Frontier Biosciences(大阪大学前沿生命科学研究生院)
;
The University of Tokyo, Faculty of Engineering(东京大学工学部)
;
The University of Osaka, Graduate School of Engineering Science(大阪大学工学研究院)
;
The University of Osaka, Graduate School of Medicine(大阪大学医学研究院)