机构
*
Tsinghua University(清华大学)
;
Fudan University(复旦大学)
;
City University of Hong Kong(香港城市大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
MPIB: A Benchmark for Medical Prompt Injection Attacks and Clinical Safety in LLMs
MPIB:用于医疗提示注入攻击和大语言模型临床安全性的基准
Junhyeok Lee, Han Jang, Kyu Sung Choi
机构
*
Seoul National University College of Medicine(首尔国立大学医学院)
;
Seoul National University(首尔国立大学)
;
Seoul National University College of Medicine, Seoul National University Hospital(首尔国立大学医学院、首尔国立大学医院)
CASTLE: A Comprehensive Benchmark for Evaluating Student-Tailored Personalized Safety in Large Language Models
CASTLE:一个评估大语言模型学生定制个性化安全性的综合基准
Rui Jia, Ruiyi Lan, Fengrui Liu, Zhongxiang Dai, Bo Jiang, Jing Shao, Jingyuan Chen, Guandong Xu, Fei Wu, Min Zhang
机构
*
East China Normal University(华东师范大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Zhejiang University(浙江大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Education University of Hong Kong(香港教育大学)
机构
*
Peking University(北京大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
City University of Hong Kong(香港城市大学)
;
Fudan University(复旦大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Tsinghua University(清华大学)
;
Zhejiang University(浙江大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学香槟分校)
;
Marquette University(马凯特大学)
;
Juniata College(朱尼阿特学院)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
大语言模型还能自我解释吗?探讨量化对自我解释的影响
Qianli Wang, Nils Feldhus, Pepa Atanasova, Fedor Splitt, Simon Ostermann, Sebastian Möller, Vera Schmitt
机构
*
Quality and Usability Lab, Technische Universität Berlin(柏林技术大学质量与可用性实验室)
;
University of Copenhagen(哥本哈根大学)
;
Saarland Informatics Campus(萨尔州信息学校园)
;
German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)
;
Centre for European Research in Trusted AI (CERTAIN)(可信AI欧洲研究中心)
;
BIFOLD – Berlin Institute for the Foundations of Learning and Data(柏林学习与数据基础研究院)
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Baidu Inc.(百度公司)