From Punchlines to Predictions: A Metric to Assess LLM Performance in Identifying Humor in Stand-Up Comedy
从笑点到预测:一种评估大语言模型在识别单口喜剧幽默能力的指标
机构 * International Christian University(国际基督教大学) ; University of Tsukuba(茨口大学)
AI总结 本研究提出了一种评估大语言模型识别单口喜剧幽默能力的指标,发现即使领先模型在幽默检测上的表现也仅达到51%,低于人类的41%。
Comments Accepted to CMCL2025 @ NAACL