Adversarial Defence without Adversarial Defence: Enhancing Language Model Robustness via Instance-level Principal Component Removal
机构 * The University of Manchester, UK(曼彻斯特大学) ; Durham University, UK(杜伦大学) ; The University of Southampton, UK(南安普顿大学) ; Automated Analytics, UK(自动化分析)
Comments This paper was accepted with an A-decision to Transactions of the Association for Computational Linguistics. This version is the pre-publication version prior to MIT Press production