LZ Penalty: An information-theoretic repetition penalty for autoregressive language models
LZ惩罚:一种用于自回归语言模型的信息论重复惩罚项
机构 * Salesforce AI Research(Salesforce人工智能研究)
专题命中 推理与问题求解 :language model(title,abstract);分类 cs.AI、cs.LG
AI总结 本文提出LZ惩罚,基于LZ77压缩算法,可让开源推理模型用贪婪解码时无退化重复,优于现有频率、重复惩罚。
Comments Post-publication corrections (minor calculation mistakes)