From Coarse to Fine: Recursive Audio-Visual Semantic Enhancement for Speech Separation
机构 * School of Cyberspace Science and Technology(网络空间科学与技术学院) ; Beijing Institute of Technology(北京理工大学) ; Sun Yat-sen University(中山大学) ; Qilu University of Technology(齐鲁工业大学) ; Shandong Computer Science Center(山东计算机科学中心) ; School of Information and Electronics(信息电子学院)
专题命中 音频语音多模态 :audio-visual(title,abstract)