机构
*
Queen Mary University of London(女王玛丽大学)
;
Shanghai Sixth People’s Hospital Affiliated to SJTU School of Medicine(上海第六人民医院附属复旦大学医学院)
;
University College London(大学学院伦敦)
CommentsAccepted non-archival paper at the CVPR 2026 AUTOPILOT Workshop (Autonomous Understanding Through Open-world Perception and Integrated Language Models for On-road Tasks)
UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation
UCSF-PDGM-VQA: 用于脑肿瘤MRI解读的视觉问答数据集
Shiv Ghosh, Junayd Lateef, Chih-Hua Liu, Yannan Yu, Andreas M. Rauschecker, Madhumita Sushil
机构
*
Fung Institute for Engineering Leadership(工程领导力基金会)
;
University of California, Berkeley(加州大学伯克利分校)
;
Department of Radiology(放射科)
;
University of California, San Francisco(加州大学旧金山分校)
;
Division of Clinical Informatics and Digital Transformation(临床信息学与数字转型部)
;
Department of Neurological Surgery(神经外科部)
Comments9 pages, 1 figure. Position paper with simulated experimental analysis of AI reliability in medication decision systems. Minor Correction to Title Metadata (Typo Fix)
Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance
位置:让我们开发数据探针,以根本理解数据如何影响大语言模型性能
Shiqiang Wang, Herbert Woisetschläger, Hans Arno Jacobsen, Mingyue Ji
机构
*
Department of Computer Science, University of Exeter, UK(埃克塞特大学计算机科学系)
;
Technical University of Munich, Germany(慕尼黑技术大学)
;
Department of Electrical and Computer Engineering, University of Toronto, Canada(多伦多大学电气与计算机工程系)
;
Department of Electrical and Computer Engineering, University of Florida, FL, USA(佛罗里达大学电气与计算机工程系)
QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI
QQJ: 量化定性判断以实现可扩展且与人类对齐的生成AI评估
Marjan Veysi, Pirooz Shamsinejadbabaki, Mohammad Zare, Mohammad Sabouri
机构
*
AI Lab, Arioobarzan Engineering Team(艾伊罗巴赞工程团队人工智能实验室)
;
Department of Computer Engineering and Information Technology(计算机工程与信息科技系)
;
Department of Informatics, Bioengineering, Robotics and Systems Engineering(信息学、生物工程、机器人与系统工程系)
;
University of Genoa(热那亚大学)
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
Columbia University(哥伦比亚大学)
;
California State University(加州州立大学)
;
University of Montreal(蒙特利尔大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Rensselaer Polytechnic Institute(莱斯利理工学院)
;
The University of Manchester(曼彻斯特大学)
;
Harvard University(哈佛大学)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
检索增强生成系统中的可信度:综述
Yujia Zhou, Wenbo Zhang, Jingying Shao, Yan Liu, Xiaoxi Li, Jiajie Jin, Hongjin Qian, Zheng Liu, Chaozhuo Li, Jason Chen Zhang, Zhicheng Dou, Philip S. Yu, Jiaxin Mao
机构
*
Tsinghua University(清华大学)
;
Renmin University of China(中国人民大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Hong Kong Polytechnic University(香港理工大学)
;
Microsoft Research Asia(微软亚洲研究院)
;
University of Illinois(伊利诺伊大学)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
关注人工智能:一种有效的人工智能系统人类监督的框架
Susanne Gaube, Markus Langer, Tim Miller, Kevin Baum, Raimund Dachselt, Anna Maria Feit, Ujwal Gadiraju, Harmanpreet Kaur, Mark T. Keane, Richard Landers, Johann Laux, Q. Vera Liao, Brian Lim, Linda Onnasch, Tim Schrills, Liz Sonenberg, Chenhao Tan, Nava Tintarev, Ziang Xiao, Hanwei Zhang
机构
*
University College London(伦敦大学)
;
University of Freiburg(弗赖堡大学)
;
University of Queensland(昆士兰大学)
;
Saarland University(萨尔兰大学)
;
TU Dresden(德累斯顿技术大学)
;
Delft University of Technology(代尔夫特理工大学)
;
University of Minnesota(明尼苏达大学)
;
University College Dublin(都柏林大学)
;
University of Oxford(牛津大学)
;
University of Michigan(密歇根大学)
;
National University of Singapore(新加坡国立大学)
;
Technische Universität Berlin(柏林技术大学)
;
University of Lübeck(吕贝克大学)
;
University of Melbourne(墨尔本大学)
;
University of Chicago(芝加哥大学)
;
Maastricht University(马斯特里赫特大学)
CommentsThe conceptual analysis for this work was undertaken by the authors at Dagstuhl seminar 25272 'Challenges of Human Oversight: Achieving Human Control of AI-Based Systems' (https://www.dagstuhl.de/25272), held at Schloss Dagstuhl (June 29th-July 4th, 2025)
Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes
帮助陷入困境的客户:一个基于LLM的代理,能够对话、探测和分流
Alankar Atreya, Stefan Sylvius Wanger, Devesh Batra, Robert Hankache, Cristovao Iglesias, Patrick Sinclair, Giulio Pelosio, Michael McMillan, Greig A. Cowan, Raad Khraishi