BESPOKE: Benchmark for Search-Augmented Large Language Model Personalization via Diagnostic Feedback
BESPOKE:通过诊断反馈进行搜索增强型大语言模型个性化的基准测试
机构 * Department of Artificial Intelligence, Yonsei University, Seoul, Republic of Korea(人工智能系,延世大学,首尔,大韩民国)
AI总结 提出BESPOKE基准,通过收集真实用户历史并配对细粒度偏好分数与反馈,系统评估搜索增强型大语言模型在个性化信息检索任务中的表现。
Comments Accepted to ICML 2026