机构
*
Xinjiang Multimodal Intelligent Processing and Information Security Engineering Technology Research Center, School of Computer Science and Technology, Xinjiang University(新疆多模态智能处理与信息安全工程技术创新中心,计算机科学与技术学院,新疆大学)
;
Department of Computer Science and Technology, Tsinghua University(计算机科学与技术系,清华大学)
;
School of Electrical Engineering and Automation, Tianjin University of Technology(电气工程与自动化学院,天津工业大学)
MIKU-PAL: An Automated and Standardized Multi-Modal Method for Speech Paralinguistic and Affect Labeling
Yifan Cheng, Ruoyi Zhang, Jiatong Shi
机构
*
Santa Clara, CA, USA(美国圣克拉拉)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
Nanjing University of Information Science and Technology(南京信息工程大学)
机构
*
Hangzhou Dianzi University(杭州电子科技大学)
;
Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
;
Xi’an Jiaotong University(西安交通大学)
;
Macao Polytechnic University(澳门 polytechnic university)
机构
*
Defense Innovation Institute, Academy of Military Sciences, Beijing, China(国防科技研究院,军事科学院,北京,中国)
;
Tianjin Artificial Intelligence Innovation Center, Tianjin, China(天津人工智能创新中心,天津,中国)
;
Harbin Institute of Technology, Shenzhen, China(哈尔滨工业大学,深圳,中国)
;
Tianjin University, Tianjin, China(天津大学,天津,中国)
UWAV: Uncertainty-weighted Weakly-supervised Audio-Visual Video Parsing
Yung-Hsuan Lai, Janek Ebbers, Yu-Chiang Frank Wang, François Germain, Michael Jeffrey Jones, Moitreya Chatterjee
机构
*
Graduate Institute of Communication Engineering, National Taiwan University(国立台湾大学通信工程研究所)
;
NVIDIA, Taiwan(台湾NVIDIA公司)
;
Mitsubishi Electric Research Labs (MERL)(三菱电机研究实验室)