Safe in the Future, Dangerous in the Past: Dissecting Temporal and Linguistic Vulnerabilities in LLMs
未来安全,过去危险:剖析大语言模型中的时间与语言脆弱性
机构 * African Institute for Mathematical Science(非洲数学科学研究所) ; University of Vienna(维也纳大学)
专题命中 安全训练 :alignment(abstract);safety(abstract);分类 cs.CL、cs.AI、cs.CY
AI总结 研究发现大语言模型在不同语言和时间框架下存在显著安全差异,提出不变对齐以提升跨语言和时间的稳定性。