To Memorize or to Retrieve: Scaling the Interaction Between Pretraining and Retrieval
记忆还是检索:考虑RAG的缩放规律
机构 * Stanford University(斯坦福大学) ; Independent Researcher(独立研究员) ; Patronus AI ; The Ohio State University(俄亥俄州立大学) ; Carnegie Mellon University(卡内基梅隆大学)
专题命中 预训练与数据 :pretraining(title,abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 研究探讨了预训练知识与检索知识的平衡,提出三维缩放框架,揭示在不同模型规模和任务类型下检索的边际效用,为语言模型设计提供数据资源分配指导。
Comments Code available at https://github.com/DegenAI-Labs/RAG-Scaling-Laws