Generating Realistic Safety-Critical Scenarios for Vehicle-Pedestrian Interactions
生成车辆-行人交互的安全关键场景
Qingwen Pu, Kun Xie, Yuan Zhu, Guocong Zhai
机构
*
Transportation Informatics Lab, Department of Civil and Environmental Engineering, Old Dominion University(交通信息实验室,土木与环境工程系,旧 Dominion 大学)
;
Inner Mongolia Center for Transportation Research, Inner Mongolia University(内蒙古交通研究所,内蒙古大学)
;
School of Transportation and Logistics, National Engineering Laboratory of Integrated Transportation Big Data Application Technology, National and Local Joint Engineering Research Center of Integrated Transportation Intelligence, Southwest Jiaotong University(交通运输学院,国家集成交通大数据应用技术工程实验室,国家与地方联合集成交通智能工程研究中心,西南交通大学)
Unlearning Isn't Deletion: Investigating Reversibility of Machine Unlearning in LLMs
反学习不是删除:调查机器反学习在大语言模型中的可逆性
Xiaoyu Xu, Xiang Yue, Yang Liu, Qingqing Ye, Huadi Zheng, Peizhao Hu, Minxin Du, Haibo Hu
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of California, Santa Cruz(加州大学圣克ruz分校)
;
Huawei Technologies(华为技术有限公司)
;
Research Centre for Privacy and Security Technologies in Future Smart Systems, PolyU(未来智能系统中的隐私与安全技术研究中心,PolyU)
Prioritizing High-Consequence Biological Capabilities in Evaluations of Artificial Intelligence Models
在评估人工智能模型时优先考虑高后果生物能力
Jaspreet Pannu, Doni Bloomfield, Alex Zhu, Robert MacKnight, Gabe Gomes, Anita Cicero, Thomas V. Inglesby
机构
*
Center for Health Security, Bloomberg School of Public Health, Johns Hopkins University(健康安全中心,公共卫生学院,约翰霍普金斯大学)
;
Department of Health Policy, Stanford School of Medicine, Stanford University(健康政策系,斯坦福医学院,斯坦福大学)
;
Department of Chemical Engineering, Carnegie Mellon University(化学工程系,卡内基梅隆大学)
;
Department of Chemistry, Carnegie Mellon University(化学系,卡内基梅隆大学)
;
Wilton E. Scott Institute for Energy Innovation, Carnegie Mellon University(威尔顿·E·斯科特能源创新研究所,卡内基梅隆大学)
REBAR: Reference Ethical Benchmark for Autonomy Readiness
REBAR:自主性准备的参考伦理基准
Jonathan Diller, David Barnes, Rebekah Bogdanoff, Rhett Collier, Roddy Collins, Keith Fieldhouse, Yonatan Gefen, Cameron Johnson, Anuriha Kodali, Brad Kriel, Varun Murali, James Niehaus, Mish Sukharev, Joseph VanPelt, Anthony Hoogs, Vijay Kumar, Arslan Basharat
机构
*
University of Pennsylvania(宾夕法尼亚大学)
;
David Barnes, LLC(大卫·巴恩斯公司)
;
Kitware, Inc.(Kitware公司)
;
Duality Robotics, Inc.(Duality机器人公司)
;
Texas A&M University(德克萨斯大学)
;
Charles River Analytics(查尔斯河分析公司)
机构
*
Southern University of Science and Technology(南方科技大学)
;
University of Science and Technology of China(中国科学技术大学)
;
University of Birmingham(伯明翰大学)
;
Zhejiang University(浙江大学)
;
East China Normal University(华东师范大学)
;
Alibaba Group(阿里巴巴集团)
Adversarial Fragility and Language Vulnerability in Clinical AI: A Systematic Audit of Diagnostic Collapse Under Imperceptible Perturbations and Cross-Lingual Drift in Low-Resource Healthcare Settings
QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI
QQJ: 量化定性判断以实现可扩展且与人类对齐的生成AI评估
Marjan Veysi, Pirooz Shamsinejadbabaki, Mohammad Zare, Mohammad Sabouri
机构
*
AI Lab, Arioobarzan Engineering Team(艾伊罗巴赞工程团队人工智能实验室)
;
Department of Computer Engineering and Information Technology(计算机工程与信息科技系)
;
Department of Informatics, Bioengineering, Robotics and Systems Engineering(信息学、生物工程、机器人与系统工程系)
;
University of Genoa(热那亚大学)
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
Columbia University(哥伦比亚大学)
;
California State University(加州州立大学)
;
University of Montreal(蒙特利尔大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Rensselaer Polytechnic Institute(莱斯利理工学院)
;
The University of Manchester(曼彻斯特大学)
;
Harvard University(哈佛大学)
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey
检索增强生成系统中的可信度:综述
Yujia Zhou, Wenbo Zhang, Jingying Shao, Yan Liu, Xiaoxi Li, Jiajie Jin, Hongjin Qian, Zheng Liu, Chaozhuo Li, Jason Chen Zhang, Zhicheng Dou, Philip S. Yu, Jiaxin Mao
机构
*
Tsinghua University(清华大学)
;
Renmin University of China(中国人民大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Hong Kong Polytechnic University(香港理工大学)
;
Microsoft Research Asia(微软亚洲研究院)
;
University of Illinois(伊利诺伊大学)
Keeping an Eye on AI: A Framework for Effective Human Oversight of AI Systems
关注人工智能:一种有效的人工智能系统人类监督的框架
Susanne Gaube, Markus Langer, Tim Miller, Kevin Baum, Raimund Dachselt, Anna Maria Feit, Ujwal Gadiraju, Harmanpreet Kaur, Mark T. Keane, Richard Landers, Johann Laux, Q. Vera Liao, Brian Lim, Linda Onnasch, Tim Schrills, Liz Sonenberg, Chenhao Tan, Nava Tintarev, Ziang Xiao, Hanwei Zhang
机构
*
University College London(伦敦大学)
;
University of Freiburg(弗赖堡大学)
;
University of Queensland(昆士兰大学)
;
Saarland University(萨尔兰大学)
;
TU Dresden(德累斯顿技术大学)
;
Delft University of Technology(代尔夫特理工大学)
;
University of Minnesota(明尼苏达大学)
;
University College Dublin(都柏林大学)
;
University of Oxford(牛津大学)
;
University of Michigan(密歇根大学)
;
National University of Singapore(新加坡国立大学)
;
Technische Universität Berlin(柏林技术大学)
;
University of Lübeck(吕贝克大学)
;
University of Melbourne(墨尔本大学)
;
University of Chicago(芝加哥大学)
;
Maastricht University(马斯特里赫特大学)
CommentsThe conceptual analysis for this work was undertaken by the authors at Dagstuhl seminar 25272 'Challenges of Human Oversight: Achieving Human Control of AI-Based Systems' (https://www.dagstuhl.de/25272), held at Schloss Dagstuhl (June 29th-July 4th, 2025)
Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes
帮助陷入困境的客户:一个基于LLM的代理,能够对话、探测和分流
Alankar Atreya, Stefan Sylvius Wanger, Devesh Batra, Robert Hankache, Cristovao Iglesias, Patrick Sinclair, Giulio Pelosio, Michael McMillan, Greig A. Cowan, Raad Khraishi
Ancient Greek to Modern Greek Machine Translation: A Novel Benchmark and Fine-Tuning Experiments on LLMs and NMT Models
古希腊语到现代希腊语机器翻译:一种新的基准和对LLM和NMT模型的微调实验
Spyridon Mavromatis, Sokratis Sofianopoulos, Prokopis Prokopidis, Maria Giagkou
机构
*
National and Kapodistrian University of Athens, Department of Informatics and Telecommunications(雅典国家和卡普迪斯特里亚大学信息与电信系)
;
Institute for Language and Speech Processing, Athena RC(语言与语音处理研究所,雅典RC)
Comments14 pages. Accepted for presentation at the 15th Language Resources and Evaluation Conference (LREC 2026), Palma, Mallorca, Spain
Journal refProceedings of the Fifteenth Language Resources and Evaluation Conference (LREC 2026), pp. 8685-8698. European Language Resources Association (ELRA)
机构
*
Department of Software Engineering(软件工程系)
;
Blekinge Institute of Technology(布莱金厄理工大学)
;
Department of Computer Science(计算机科学系)
;
Malmö University(马尔默大学)
;
Al-Balqa Applied University(阿尔巴卡应用大学)
Conv-FinRe: A Conversational and Longitudinal Benchmark for Utility-Grounded Financial Recommendation
Conv-FinRe:一种用于实用导向财务推荐的对话和纵向基准
Yan Wang, Yi Han, Lingfei Qian, Yueru He, Xueqing Peng, Dongji Feng, Zhuohan Xie, Vincent Jim Zhang, Rosie Guo, Fengran Mo, Jimin Huang, Yankai Chen, Xue Liu, Jian-Yun Nie
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
Columbia University(哥伦比亚大学)
;
California State University(加州州立大学)
;
University of Montreal(蒙特利尔大学)
;
The University of Manchester(曼彻斯特大学)
;
McGill University(麦吉尔大学)