机构
*
School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学)
;
Key Laboratory of Machine Intelligence and System Control, Shandong University(机器智能与系统控制重点实验室,山东大学)
;
Department of Medicine and Therapeutics, The Chinese University of Hong Kong(医学与治疗学系,香港中文大学)
;
Department of Geriatric Medicine, Qilu Hospital of Shandong University(老年医学科,山东大学齐鲁医院)
;
Department of Psychiatry, The Chinese University of Hong Kong(精神病学系,香港中文大学)
;
Li Chiu Kong Family Sleep Assessment Unit, Department of Psychiatry, Faculty of Medicine, The Chinese University of Hong Kong(李秋虹家庭睡眠评估单元,精神病学系,医学院,香港中文大学)
;
Li Ka Shing Institute of Health Sciences, Faculty of Medicine, The Chinese University of Hong Kong(李嘉诚健康科学研究院,医学院,香港中文大学)
;
Gerald Choa Neuroscience Institute, Department of Medicine and Therapeutics, The Chinese University of Hong Kong(Gerald Choa 神经科学研究所,医学与治疗学系,香港中文大学)
机构
*
Northeastern University(东北大学)
;
University of Notre Dame(Notre Dame 大学)
;
University of Waterloo(滑铁卢大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Adobe(Adobe公司)
;
Microsoft Research Asia(微软亚洲研究院)
PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects
PolySpeech-100:面向100多种语言和方言的大规模语音理解基准
Sicheng Yang, Shulan Ruan, Shiwei Wu, Yu Liu, Lu Fan, Zhi Li, You He
机构
*
Shenzhen International Graduate School, Tsinghua University(深圳国际研究生院,清华大学)
;
Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)
;
JD AI Research(京东人工智能研究院)
机构
*
Department of Computer Science and Engineering(计算机科学与工程系)
;
Department of Electrical and Electronic Engineering(电气与电子工程系)
;
Islamic University of Technology(伊斯兰技术大学)
An Empirical Evaluation of LLM-Generated Code Security Across Prompting Methods
LLM生成代码安全性的提示方法实证评估
Mohammed Kharma, Ahmed Sabbah, Mohammad Alkhanafseh, Mohammad Hammoudeh, David Mohaisen
机构
*
Department of Computer Science, Birzeit University(计算机科学系,巴勒斯坦比泽大学)
;
King Fahd University of Petroleum and Minerals(国王法赫德石油和矿物大学)
;
University of Central Florida(中央佛罗里达大学)
TRIP-Evaluate: An Open Multimodal Benchmark for Evaluating Large Models in Transportation
TRIP-Evaluate: 一个用于评估交通领域大模型的开放多模态基准
Han Gong, Zhen Zhou, Yunyang Shi, Yan Tan, Jinbiao Huo, Qi Hong, Zhiyuan Liu
机构
*
School of Transportation(交通学院)
;
Southeast University(东南大学)
;
School of Artificial Intelligence and Computer Science(人工智能与计算机科学学院)
;
Jiangnan University(江南大学)
;
Department of Civil and Environmental Engineering(土木与环境工程系)
;
Hong Kong Polytechnic University(香港理工大学)
Making AI-Assisted Grant Evaluation Auditable without Exposing the Model
在不暴露模型的情况下使AI辅助的资助评估可审计
Kemal Bicakci
机构
*
Informatics Institute, Istanbul Technical University, Istanbul, Türkiye(伊斯坦布尔技术大学信息学院,伊斯坦布尔,土耳其)
;
Securify Information Technology and Security Training Consulting Inc., Ankara, Türkiye(Securify信息科技与安全培训咨询公司,安卡拉,土耳其)
Navigating Large-Scale Document Collections: MuDABench for Multi-Document Analytical QA
在大规模文档集合中导航:MuDABench用于多文档分析问答
Zhanli Li, Yixuan Cao, Lvzhou Luo, Ping Luo
机构
*
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences (CAS)(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Wenlan School of Business, Zhongnan University of Economics and Law(中南财经政法大学文澜商学院)
CommentsFindings of ACL 2026. The camera-ready version corrects some labeling errors. The accompanying repository is continuously updated based on community feedback; for the most up-to-date implementation and results, please refer to the repository
RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents
RPA-Check:一种多阶段自动框架,用于评估基于大语言模型的角色扮演代理
Riccardo Rosati, Edoardo Colucci, Massimiliano Bolognini, Adriano Mancini, Paolo Sernani
机构
*
Department of Political Sciences, Communication and International Relations, University of Macerata(马切拉塔大学政治学、传播与国际关系系)
;
Department of Law, University of Macerata(马切拉塔大学法律系)
EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration
EVGeoQA:基于动态多目标地理空间探索的LLM基准测试
Jianfei Wu, Zhichun Wang, Zhensheng Wang, Zhiyu He
机构
*
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
;
Beijing Key Laboratory of Artificial Intelligence for Education(北京市教育人工智能重点实验室)
;
Engineering Research Center of Intelligent Technology and Educational Application, Ministry of Education(教育部智能技术与教育应用工程研究中心)
;
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机学院)
Pan Chen, Shaohong Chen, Mark Wang, Shi Xuan Leong, Priscilla Fung, Varinia Bernales, Alan Aspuru-Guzik
机构
*
University of Toronto(多伦多大学)
;
Nanyang Technological University(南洋理工大学)
;
Acceleration Consortium(加速联盟)
;
Vector Institute for Artificial Intelligence(向量人工智能研究所)
;
Canadian Institute for Advanced Research (CIFAR)(加拿大高等研究院(CIFAR))
;
NVIDIA(英伟达)
机构
*
Lehigh University(里海大学)
;
Squirrel Ai Learning(松鼠AI学习)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Michigan State University(密歇根州立大学)