Putting a Face to Forgetting: Continual Learning meets Mechanistic Interpretability
为遗忘赋予面孔:持续学习与机制可解释性相遇
机构 * KU Leuven(根特大学) ; University of Groningen(Groningen大学) ; Amazon(亚马逊)
AI总结 本文提出一个机制框架,从几何角度解释持续学习中的灾难性遗忘,通过玩具模型分析并验证了深度对遗忘的影响,展示了如何利用Crosscoders分析实际模型。
Comments To appear in the Proceedings of the Fifth Conference on Lifelong Learning Agents (CoLLAs), 2026