Do LLMs Need Inherent Reasoning Before Reinforcement Learning? A Study in Korean Self-Correction
LLMs是否需要在强化学习前具备内在推理能力?韩国自我修正研究
机构 * ETRI
专题命中 复杂问题求解 :reasoning(title,abstract);self-correction(title,abstract);分类 cs.CL、cs.AI
AI总结 本研究探讨LLMs在强化学习前是否需要具备内在推理能力,通过韩国自我修正研究发现,对齐模型内部推理与韩语输入关键,提升多语言推理效果。
Comments IJCNLP-AACL 2025 (Main), Outstanding Paper Award