The Mental World of Large Language Models in Recommendation: A Benchmark on Association, Personalization, and Knowledgeability
大语言模型在推荐系统中的心理世界:关联性、个性化与知识性基准测试
AI总结 本文提出LRWorld基准,评估大语言模型在推荐系统中的关联性、个性化和知识性能力,发现其在浅层相似性任务上表现良好,但在深度个性化嵌入和多模态推理方面仍有不足。
Comments 21 pages, 13 figures, 27 tables, submission to KDD 2025