OmniAssistBench: Assistant-style Interaction Benchmark for Omni-LLMs
OmniAssistBench:面向全模态大语言模型(Omni-LLMs)的助手式交互基准
机构 * Nankai University(南开大学) ; University of Waterloo(滑铁卢大学) ; Nanjing University(南京大学)
专题命中 视频数据与评测 :video understanding(abstract);分类 cs.CV
AI总结 针对全模态大语言模型助手式交互的评估瓶颈,构建OmniAssistBench数据集,实验显示Gemini-3-Pro、Qwen3-Omni-Instruct表现存在差距,当前模型在多轮交互等方面仍有不足。
Comments Project page: this https URL (https://xianyunsun.github.io/OmniAssistBench/)