Enhancing Speech Emotion Recognition Leveraging Aligning Timestamps of ASR Transcripts and Speaker Diarization
机构 * Department of English(英语系) ; National Taiwan Normal University(台湾师范大学) ; Department of Computer Science and Information Engineering(计算机科学与信息工程系)
专题命中 音频语音多模态 :multimodal(abstract);分类 cs.CL、eess.AS
Comments 6 pages, 3 figures, to appear in the Proceedings of the 2025 International Conference on Asian Language Processing (IALP)