机构
*
University of Rochester(罗切斯特大学)
;
National University of Singapore(新加坡国立大学)
;
SJTU Paris Elite Institute of Technology(上海科技技术巴黎精英学院)
;
Singapore Management University(新加坡管理学院)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
Can LLM Safety Be Ensured by Constraining Parameter Regions?
通过约束参数区域能否确保大语言模型的安全性?
Zongmin Li, Jian Su, Farah Benamara, Aixin Sun
机构
*
Nanyang Technological University(南洋理工大学)
;
Institute for Infocomm Research (I 2 R)(信息通信研究所)
;
IRIT, Université de Toulouse, CNRS, Toulouse INP(图卢兹大学IRIT研究所、法国国家科学研究中心)
;
IPAL, CNRS-NUS-A*STAR(IPAL、法国国家科学研究中心-南洋理工大学-A*STAR)
机构
*
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
;
Institute of Artificial Intelligence, Beihang University(北京航空航天大学人工智能研究院)
;
School of Software, Beihang University(北京航空航天大学软件学院)
;
State Key Laboratory of Media Convergence and Communication, Communication University of China(中国传媒大学媒体融合与传播国家重点实验室)
;
Department of Computer Science, National University of Singapore(新加坡国立大学计算机科学系)
;
Control Science and Engineering, Shandong University(山东大学控制科学与工程学院)
;
Artificial Intelligence Research Center, Lobachevsky State University(洛夫奇夫斯基国立大学人工智能研究中心)