SelfAug: Mitigating Catastrophic Forgetting in Retrieval-Augmented Generation via Distribution Self-Alignment
机构 * University of Science and Technology of China(中国科学技术大学) ; Xiaohongshu Inc.(小红书公司)
专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * University of Science and Technology of China(中国科学技术大学) ; Xiaohongshu Inc.(小红书公司)
专题命中 其他安全 :alignment(title,abstract);分类 cs.CL、cs.AI
机构 * Northeastern University(东北大学)
专题命中 其他安全 :alignment(abstract);分类 cs.CL、cs.AI
机构 * NVIDIA ; University of Toronto(多伦多大学) ; Vector Institute(向量研究所)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Comments Project page: https://research.nvidia.com/labs/toronto-ai/LuxDiT/
专题命中 其他安全 :alignment(abstract);分类 cs.LG
Comments Reinforcement Learning Conference 2025
机构 * University of Southhampton(南安普顿大学) ; KAIST(韩国科学技术院)
专题命中 其他安全 :alignment(abstract);分类 cs.AI
Comments 17 pages, 5 figures, 9 tables
机构 * Department of Informatics, King's College London(伦敦国王学院信息学院) ; Department of Engineering, King's College London(伦敦国王学院工程学院) ; John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院)
专题命中 其他安全 :alignment(abstract)
专题命中 其他安全 :safety(abstract)
专题命中 其他安全 :safety(abstract)
专题命中 其他安全 :alignment(abstract)