The Impact of Quantization on Large Reasoning Model Reinforcement Learning
量化对大推理模型强化学习的影响
机构 * Pennsylvania State University(宾夕法尼亚州立大学) ; d-Matrix
专题命中 数学推理 :reasoning(title,abstract);分类 cs.LG
AI总结 本文研究了量化对大推理模型强化学习的影响,发现量化感知训练对学习过程有负面影响,而PTQ和QLoRA能提升性能。
Comments Accepted to the NeurIPS 2025 Efficient Reasoning Workshop