PAL: Probing Audio Encoders via LLMs -- Audio Information Transfer into LLMs
PAL: 通过LLM探测音频编码器——音频信息转移到LLM中
机构 * Centre for Vision, Speech and Signal Processing (CVSSP), University of Surrey, UK(视觉、语音和信号处理中心(CVSSP),英国萨里大学) ; Surrey Institute for People-Centred AI, University of Surrey, Guildford, GU2 7XH, UK(萨里大学以人为中心的人工智能研究所,英国Guildford) ; Mohamed bin Zayed University of Artificial Intelligence (MBZUAI), Abu Dhabi, UAE(穆罕默德·本·扎耶德人工智能大学(MBZUAI),阿布扎赫德)
专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
AI总结 PAL通过LLM高效探测音频编码器,利用LAL和PLITS结合的方法提升音频信息整合效率,降低计算和内存开销。
Comments 20 pages, 7 figures