Adaptive Defense against Harmful Fine-Tuning for Large Language Models via Bayesian Data Scheduler
Zixuan Hu, Li Shen, Zhenyi Wang, Yongxian Wei, Dacheng Tao
机构
*
Nanyang Technological University(南洋理工大学)
;
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
;
University of Central Florida(佛罗里达大学)
;
Tsinghua University(清华大学)
Subhabrata Majumdar, Brian Pendleton, Abhishek Gupta
机构
*
Vijil / AI Risk and Vulnerability Alliance(Vijil / AI风险与漏洞联盟)
;
AI Risk and Vulnerability Alliance(AI风险与漏洞联盟)
;
Montreal AI Ethics Institute(蒙特利尔人工智能伦理研究所)
SC-LoRA: Balancing Efficient Fine-tuning and Knowledge Preservation via Subspace-Constrained LoRA
Minrui Luo, Fuhang Kuang, Yu Wang, Zirui Liu, Tianxing He
机构
*
Shanghai Qi Zhi Institute(上海启智研究院)
;
Institute for Interdisciplinary Information Sciences, Tsinghua University(清华大学交叉信息研究院)
;
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
Xiongan AI Institute(雄安人工智能研究院)
Approximate Fiber Products of Schemes and Their Étale Homotopical Invariants
Dongfang Zhao
专题命中
其他安全
:alignment(abstract)
CommentsSeveral experts pointed out technical flaws of this work, for example the incorrect notations being used in Section 3 and the weak connection to the claim LLM applications in Section 1. We think it is best to be withdrawn at this point so that readers will not be misled