机构
*
Tianmushan Laboratory, Beihang University(北航之梦实验室,北京航空航天大学)
;
Hangzhou International Innovation Institute, Beihang University(杭州国际创新研究院,北京航空航天大学)
;
Foshan Graduate School of Innovation, Northeastern University(佛山创新研究生学院,东北大学)
;
School of Aeronautic Science and Engineering, Beihang University(航空科学与工程学院,北京航空航天大学)
;
Faculty of Robot Science and Engineering, Northeastern University(机器人科学与工程学院,东北大学)
机构
*
Department of Computer and Network Engineering(计算机与网络工程系)
;
United Arab Emirates University(阿联酋大学)
;
Technology Innovation Institute(技术创新研究院)
;
Eötvös Loránd University(埃斯特哈齐·洛朗大学)
;
Research Institute for Digital Future(数字未来研究院)
;
Khalifa University(卡塔尔大学)
VisioMath: Benchmarking Figure-based Mathematical Reasoning in LMMs
VisioMath: 评估LMMs中基于图的数学推理能力的基准测试
Can Li, Ying Liu, Ting Zhang, Mei Wang, Hua Huang
机构
*
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
;
Beijing Key Laboratory of Artificial Intelligence for Education(北京人工智能教育重点实验室)
;
Engineering Research Center of Intelligent Technology and Educational Application, Ministry of Education(教育部智能技术与教育应用工程研究中心)
机构
*
New Laboratory of Pattern Recognition (NLPR), State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), Institute of Automation, Chinese Academy of Sciences (CASIA)(模式识别新实验室、多模态人工智能系统国家重点实验室、自动化研究所、中国科学院(CASIA))
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Peking University(北京大学)
;
Wuhan University(武汉大学)
;
ByteDance(字节跳动)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.CV、cs.AI
机构
*
Arizona State University(亚利桑那州立大学)
;
Morgan Stanley(摩根大通)
;
Rice University(里奇大学)
;
Clemson University(克莱姆森大学)
;
Northwestern University(西北大学)
;
UC Davis(加州大学戴维斯分校)
;
Iowa State University(爱荷华州立大学)
;
Washington University in St. Louis(圣路易斯华盛顿大学)
;
Columbia University(哥伦比亚大学)
;
Halmstad University(哈马格大学)
机构
*
School of Information Engineering, Minzu University of China(中国民族大学信息工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
National Library of China, Beijing, China(中国国家图书馆)
;
School of Cyberspace Security, Beijing University of Posts and Telecommunications(北京邮电大学网络安全学院)
MMR-Life: Piecing Together Real-life Scenes for Multimodal Multi-image Reasoning
MMR-Life: 组合真实场景以进行多模态多图像推理
Jiachun Li, Shaoping Huang, Zhuoran Jin, Chenlong Zhang, Pengfei Cao, Yubo Chen, Kang Liu, Jun Zhao
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
The Key Laboratory of Cognition and Decision Intelligence for Complex Systems, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所认知与决策智能复杂系统重点实验室)
;
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学交叉学科学院)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.CV、cs.AI
机构
*
School of Computer Science, University of Nottingham Ningbo China(诺丁汉大学宁波校区计算机科学学院)
;
Institute of Biomedical Engineering, Ningbo institute of materials technology and engineering, Chinese Academy of Sciences(中国科学院宁波材料技术与工程研究所生物医学工程研究所)
;
School of Electrical & Electronic Engineering, Nanyang Technological University(南洋理工大学电子与电气工程学院)
Zhibin Lan, Liqiang Niu, Fandong Meng, Jie Zhou, Jinsong Su
机构
*
School of Informatics, Xiamen University, China(厦门大学信息学院)
;
WeChat AI, Tencent Inc, China(腾讯公司微信AI部门)
;
Key Laboratory of Digital Protection and Intelligent Processing of Intangible Cultural Heritage of Fujian and Taiwan (Xiamen University), Ministry of Culture and Tourism, China(福建省和台湾非物质文化遗产数字化保护与智能处理重点实验室)
;
Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.AI、cs.LG