Learning to Interpret Weight Differences in Language Models
学习解释语言模型中的权重差异
机构 * Massachusetts Institute of Technology(麻省理工学院)
专题命中 指令微调 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 本文提出Diff Interpretation Tuning方法,通过合成标签权重差异训练适配器,使模型能用自然语言描述微调后的修改。
Comments Project code and links to weight diffs, adapters, and training data can be found at https://github.com/Aviously/diff-interpretation-tuning