ConceptGuard: Benchmarking Context-Sensitive Unlearning in Large Language Models
ConceptGuard:评估大语言模型中上下文敏感的遗忘能力
机构 * Pune Institute of Computer Technology(浦那计算机技术学院) ; University of California, Irvine(加州大学欧文分校)
AI总结 该研究针对大语言模型遗忘能力评估的缺陷,提出ConceptGuard基准,聚焦两用概念,发现现有遗忘技术存在性能缺陷,为实用安全的遗忘方法提供思路。
Comments Submitted to NeurIPS E&D Track 2026; 17 pages, 9 figures