Warm Up Before You Train: Unlocking General Reasoning in Resource-Constrained Settings
在训练前热身:在资源受限环境下解锁通用推理
机构 * Department of Computer Science, New York University Abu Dhabi(纽约大学阿布扎克分校计算机科学系)
AI总结 本文提出一种分两阶段的训练策略,在资源受限环境下通过热身提升大语言模型的推理能力。
Comments Accepted to EMNLP 2025
Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)