When Long Helps Short: How Context Length in Supervised Fine-tuning Affects Behavior of Large Language Models
机构 * X-LANCE Lab(X-LANCE实验室) ; MoE Key Lab of Artificial Intelligence(人工智能MoE关键实验室) ; AI Institute(人工智能研究院) ; School of Computer Science(计算机科学学院) ; Shanghai Jiao Tong University(上海交通大学) ; Shanghai Innovation Institution(上海创新机构) ; Jiangsu Key Lab of Language Computing(江苏语言计算关键实验室) ; Suzhou Laboratory(苏州实验室)
专题命中 指令微调 :large language model(title,abstract);language model(title,abstract);LLM(abstract);pretraining(abstract)