RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
RAVENEA:多模态检索增强视觉文化理解的基准
Jiaang Li, Yifei Yuan, Wenyan Li, Mohammad Aliannejadi, Daniel Hershcovich, Anders Søgaard, Ivan Vulić, Wenxuan Zhang, Paul Pu Liang, Yang Deng, Serge Belongie
机构
*
University of Copenhagen(哥本哈根大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Amsterdam(阿姆斯特丹大学)
;
University of Cambridge(剑桥大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Singapore University of Technology and Design(新加坡科技设计大学)
;
Singapore Management University(新加坡管理大学)
机构
*
College of Health Science and Technology, Shanghai Jiao Tong University School of Medicine, Shanghai, China(上海交通大学医学院健康科学与技术学院)
;
Fudan University, Shanghai, China(复旦大学)
;
Shanghai Innovation Institute, Shanghai, China(上海创新研究院)
;
Tsinghua University, Beijing, China(清华大学)
;
Shanghai Artificial Intelligence Laboratory, Shanghai, China(上海人工智能实验室)
;
The Chinese University of Hong Kong, Hong Kong, China(香港中文大学)
;
Ruijin Hospital, Shanghai Jiaotong University, Shanghai, China(上海交通大学瑞金医院)
专题命中
视觉推理
:grounding(abstract);multimodal large language model(abstract);分类 cs.CV、cs.AI
C^2ROPE: Causal Continuous Rotary Positional Encoding for 3D Large Multimodal-Models Reasoning
C^2ROPE: 3D 大多模态模型推理中的因果连续旋转位置编码
Guanting Ye, Qiyan Zhao, Wenhao Yu, Xiaofeng Zhang, Jianmin Ji, Yanyong Zhang, Ka-Veng Yuen
机构
*
State Key Laboratory of Internet of Things for Smart City, University of Macau(物联网智能城市国家重点实验室,澳门大学)
;
Department of Automation, Shanghai Jiaotong University(上海交通大学自动化系)
;
Institute of Advanced Technology, University of Science and Technology of China(中国科学技术大学先进技术研究院)
;
School of Computer Science and Technology, USTC(中国科学技术大学计算机科学与技术学院)
;
School of Artificial Intelligence and Data Science, USTC(中国科学技术大学人工智能与数据科学学院)
机构
*
Huazhong University of Science and Technology(华中科技大学)
;
Nanyang Technological University(南洋理工大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Renmin University of China(中国人民大学)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.AI
Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding
文本优先于视觉:针对超高清遥感理解的代理强化学习在超高清遥感理解中的知识注入至关重要
Fengxiang Wang, Mingshuo Chen, Yueying Li, Yajie Yang, Yuhao Zhou, Di Wang, Yifan Zhang, Haoyu Wang, Haiyan Zhao, Hongda Sun, Long Lan, Jun Song, Yulin Wang, Jing Zhang, Wenlong Zhang, Bo Du
机构
*
National University of Defense Technology, China(国防科技大学)
;
Beijing University of Posts and Telecommunications, China(北京邮电大学)
;
University of the Chinese Academy of Sciences, China(中国科学院大学)
;
Sichuan University, China(四川大学)
;
Wuhan University, China(武汉大学)
;
Chinese Academy of Science, China(中国科学院)
;
Tsinghua University, China(清华大学)
;
Shanghai Artificial Intelligence Laboratory, China(上海人工智能实验室)
;
Renmin University of China, China(中国人民大学)
S2SServiceBench: A Multimodal Benchmark for Last-Mile S2S Climate Services
S2SServiceBench:一个多模态基准用于最后一公里S2S气候服务
Chenyue Li, Wen Deng, Zhuotao Sun, Mengxi Jin, Hanzhe Cui, Han Li, Shentong Li, Man Kit Yu, Ming Long Lai, Yuhao Yang, Mengqian Lu, Binhang Yuan
机构
*
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Nanjing University of Information Science and Technology(南京信息工程大学)
;
Beijing Normal University(北京师范大学)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.LG
CommentsAn extended abstract of this article is accepted for presentation at AAMAS 2026: Olivares-Alarcos, A., Muhammad, A., Sanjaya, S., Lin, H. and Alenyà, G. (2026). Blending ontologies and language models to generate sound and natural robot explanations. In Proceedings of the International Conference on Autonomous Agents and Multiagent Systems. IFAAMAS
Picking the Right Specialist: Attentive Neural Process-based Selection of Task-Specialized Models as Tools for Agentic Healthcare Systems
选择合适的专家:基于神经过程的注意力机制用于选择任务专用模型作为智能医疗系统工具
Pramit Saha, Joshua Strong, Mohammad Alsharid, Divyanshu Mishra, J. Alison Noble
机构
*
Department of Engineering Science, University of Oxford, United Kingdom(牛津大学工程科学系)
;
Department of Computer Science, Khalifa University, Abu Dhabi, United Arab Emirates(哈利法大学计算机科学系)
Protect$^*$: Steerable Retrosynthesis through Neuro-Symbolic State Encoding
Protect$^*$: 通过神经符号状态编码实现可操控的逆合成
Shreyas Vinaya Sathyanarayana, Shah Rahil Kirankumar, Sharanabasava D. Hiremath, Bharath Ramsundar
机构
*
Deep Forest Sciences(深林科技)
;
Departament de Farmacologia, Toxicologia i Química Terapèutica, Universitat de Barcelona(巴塞罗那大学药理学、毒理学与治疗化学系)
;
Institut de Nanociència i Nanotecnologia IN2UB, Universitat de Barcelona(巴塞罗那大学纳米科学与纳米技术研究所)