TAROT: Test-driven and Capability-adaptive Curriculum Reinforcement Fine-tuning for Code Generation with Large Language Models
TAROT: 为基于大语言模型的代码生成设计的测试驱动和能力适应课程强化微调
机构 * Electronics and Telecommunications Research Institute(电信研究所) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) ; The Hong Kong University of Science and Technology(香港科技大学) ; Hugging Face ; Ant Group(蚂蚁集团)
专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG
AI总结 TAROT通过测试驱动和能力适应的课程强化微调方法,提升大语言模型生成代码的功能正确性和稳健性。
Comments The first three authors contributed equally to this work; listing order is random