Arabic Multimodal Machine Learning: Datasets, Applications, Approaches, and Challenges
机构 * Ziane Achour University(赞赞·阿赫尔大学)
专题命中 音频语音多模态 :multimodal(title,abstract);分类 cs.CL
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Ziane Achour University(赞赞·阿赫尔大学)
专题命中 音频语音多模态 :multimodal(title,abstract);分类 cs.CL
机构 * University of Science and Technology of China(中国科学技术大学) ; Hong Kong Polytechnic University(香港理工大学)
专题命中 音频语音多模态 :any-to-any(title,abstract)
机构 * Department of Computer Science and Engineering, Chandigarh University, Mohali, Punjab, India(昌迪加尔大学计算机科学与工程系,莫哈利,旁遮普,印度) ; Department of Computer Science and Engineering, Manav Rachna International Institute of Research and Studies(曼纳瓦拉国际研究与学习研究所计算机科学与工程系)
专题命中 音频语音多模态 :multi-modal(title,abstract)
专题命中 音频语音多模态 :multimodal(title,abstract)
机构 * National University of Singapore(国立新加坡大学) ; Nanyang Technological University(南洋理工大学) ; Tsinghua University(清华大学)
专题命中 音频语音多模态 :multimodal(abstract);multi-modal(abstract);分类 cs.CL、cs.AI
Comments Accepted by EMNLP 2025 Main
机构 * School of Marine Science and Technology, Northwestern Polytechnical University(海洋科学与技术学院,西北工业大学) ; Institute of Artificial Intelligence (TeleAI), China Telecom(人工智能研究所(TeleAI),中国电信) ; Research and Development Institute of Northwestern Polytechnical University in Shenzhen, China(西北工业大学深圳研发院,中国)
专题命中 音频语音多模态 :audio-visual(abstract)