Exploring Audio Cues for Enhanced Test-Time Video Model Adaptation
机构 * Artificial Intelligence Research Institute, Shenzhen MSU-BIT University and Guangdong-Hong Kong-Macao Joint Laboratory for Emotional Intelligence and Pervasive Computing(人工智能研究院、深圳MSU-BIT大学及粤港澳大湾区情感智能与普适计算联合实验室) ; School of Software Engineering, South China University of Technology(华南理工大学软件学院) ; College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院)
专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.LG
Comments 14 pages, 7 figures