Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models
机构 * Yuan-Hou(元侯)
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);grounding(abstract);分类 cs.CV
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Yuan-Hou(元侯)
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);grounding(abstract);分类 cs.CV
机构 * aThe State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences, China [1ex] bSchool of Artificial Intelligence, University of Chinese Academy of Sciences, China [1ex] cSchool of Artificial Intelligence, Beijing University of Posts ; Telecommunications, China [1ex] dKey Laboratory of Computing Power Network ; Shandong Computer Science Center, Qilu University of Technology (Shandong Academy of Sciences), China [1ex] e Shandong Provincial Key Laboratory of Computing Power Internet ; Service Computing, Shandong Fundamental Research Center for Computer Science, China
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV
Comments 27 pages, 11 figures. Accepted to Information Fusion. Final journal version: volume 126 (Part B), February 2026
Journal ref Information Fusion, 126 (Part B), February 2026, 103652
机构 * Shenzhen International Graduate School, Tsinghua University, China(清华大学深圳国际研究生院) ; Noah’s Ark Lab, Huawei, China(华为诺亚实验室)
专题命中 视觉定位与Grounding :grounding(title,abstract)
Comments More details and videos can be found at https://robo-map.github.io
机构 * College of Computer Science and Technology of Zhejiang University(浙江大学计算机科学与技术学院) ; Polytechnic Institute of Zhejiang University(浙江大学Polytechnic学院) ; Om AI Research(Om AI研究机构) ; Binjiang Research Institute of Zhejiang University(浙江大学滨江研究机构) ; School of Software Engineering of Zhejiang University(浙江大学软件工程学院) ; School of Mathematical Sciences of Zhejiang University(浙江大学数学科学学院) ; China Academy of Space Technology(中国航天科技研究院) ; University of Bristol(布里斯托大学)
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV、cs.AI
机构 * Department of Computer Sciences University of Wisconsin-Madison(计算机科学系威斯康星大学麦迪逊分校)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV、cs.AI
Comments NeurIPS 2025
机构 * Neuroelectrics(神经电医学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 2 Figures
机构 * Brown University(布朗大学)
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG