Are Multimodal Foundation Models All That Is Needed for Emofake Detection?
专题命中 音频语音多模态 :multimodal(title,abstract);multimodal foundation model(title,abstract);cross-modal(abstract);分类 eess.AS
Comments Accepted to APSIPA-ASC 2025
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 音频语音多模态 :multimodal(title,abstract);multimodal foundation model(title,abstract);cross-modal(abstract);分类 eess.AS
Comments Accepted to APSIPA-ASC 2025
机构 * Department of Computer Science(计算机科学系) ; Information Engineering, National Taiwan Normal University(信息工程,台湾正常大学)
专题命中 音频语音多模态 :multimodal(title,abstract);multimodal foundation model(title,abstract);分类 cs.CL、cs.AI
Comments Copyright 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works
机构 * University of Groningen, The Netherlands(Groningen大学,荷兰)
专题命中 音频语音多模态 :multimodal(title,abstract);cross-modal(abstract);audio-visual(abstract);分类 cs.CL、cs.MM
机构 * Columbia University(哥伦比亚大学) ; University of Washington(华盛顿大学)
专题命中 音频语音多模态 :cross-modal(title,abstract);audio-visual(abstract);分类 cs.CL、cs.AI、eess.AS
专题命中 音频语音多模态 :audio-visual(title,abstract);分类 eess.AS
Comments Accepted into Automatic Speech Recognition and Understanding- ASRU 2025
机构 * Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(电子与电气工程系,香港理工大学)
专题命中 音频语音多模态 :multimodal(abstract);MLLM(abstract);分类 eess.AS
Comments 5 pages, 2 figures