TDBench: A Benchmark for Top-Down Image Understanding with Reliability Analysis of Vision-Language Models
专题命中 视觉定位与Grounding :vision-language model(title);vision language model(abstract);grounding(abstract);分类 cs.AI、cs.LG
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :vision-language model(title);vision language model(abstract);grounding(abstract);分类 cs.AI、cs.LG
机构 * Carnegie Mellon University(卡内基梅隆大学)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV、cs.AI
Comments Project Page: https://deeptracereward.github.io/
机构 * Centre Inria d’Universite Cote d’Azur(法国国家信息与自动化技术研究所(Inria))
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Accepted to X-Sense Ego-Exo Sensing for Smart Mobility Workshop at ICCV 2025 Conference
机构 * University of Amsterdam(阿姆斯特丹大学) ; ETH Zürich(苏黎世联邦理工学院) ; University of Copenhagen(哥本哈根大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments ICML 2025 Assessing World Models Workshop; EMNLP 2025 Findings
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted at ICCV25
机构 * Intelligent Autonomous Systems Group(智能自主系统组) ; TU Darmstadt(图宾根大学) ; AICOR Insitute for Artificial Intelligence(人工智能研究所) ; University of Bremen(不莱梅大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(abstract)
Comments 20 pages, 6 figures. Magn Reson Med. 2025
机构 * The Chinese University of Hong Kong(香港中文大学) ; University of Science and Technology of China(中国科学技术大学) ; National University of Singapore(新加坡国立大学) ; Singapore Management University(新加坡管理学院) ; The Hong Kong Polytechnic University(香港理工大学)
专题命中 视觉定位与Grounding :grounding(abstract)