360° Image Perception with MLLMs: A Comprehensive Benchmark and a Training-Free Method
360度图像感知与MLLMs:一个全面的基准和无需训练的方法
机构 * GSIS, Tohoku University(东大GSIS研究所,东京东大大学) ; RIKEN AIP, Japan(日本RIKEN AIP)
专题命中 视觉问答 :visual question answering(abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV、cs.AI
AI总结 本文提出360Bench基准和Free360方法,评估MLLMs在360度图像感知中的能力,揭示其不足,并提出基于场景图的无需训练框架提升VQA性能。
Journal ref ECCV2026 (Link: https://tranhuyen1191.github.io/360Bench-Free360/)