Benchmarking Generation and Evaluation Capabilities of Large Language Models for Instruction Controllable Summarization
专题命中 安全评测 :alignment(abstract);分类 cs.CL、cs.LG
Comments NAACL 2024 Findings, GitHub Repo: https://github.com/yale-nlp/InstruSum, LLM-evaluators Leaderboard: https://huggingface.co/spaces/yale-nlp/InstruSumEval