CrossVid: A Comprehensive Benchmark for Evaluating Cross-Video Reasoning in Multimodal Large Language Models
CrossVid: 一个用于评估多模态大语言模型在跨视频推理中的综合基准
专题命中 推理评测 :reasoning(title,abstract);分类 cs.AI
AI总结 CrossVid是首个用于评估多模态大语言模型跨视频推理能力的综合基准,通过多样化的任务和数据集验证了模型在复杂视频推理任务中的表现。
Comments Accepted to AAAI 2026 (main track). For code and data, see https://github.com/chuntianli666/CrossVid