UVE: Are MLLMs Unified Evaluators for AI-Generated Videos?
UVE: MLLMs是否能作为AI生成视频的统一评估器?
机构 * State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机科学学院,北京大学) ; ByteDance Seed(字节跳动种子) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)
专题命中 其他VLM :multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
AI总结 本文探讨使用多模态大语言模型作为AI生成视频的统一评估器的可行性,并通过UVE-Bench基准测试评估了18种MLLMs的性能。