Optimizing User Profiles via Contextual Bandits for Retrieval-Augmented LLM Personalization
通过上下文老虎机优化用户资料以实现检索增强的LLM个性化
机构 * McGill University(麦吉尔大学) ; Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) ; Université de Montréal(蒙特利尔大学) ; Salesforce(Salesforce公司) ; HEC Montréal(蒙特利尔HEC商学院) ; Mila - Quebec AI Institute(魁北克人工智能研究所)
AI总结 本文提出PURPLE框架,通过上下文老虎机优化用户资料,利用Plackett-Luce模型捕捉复杂依赖,提升LLM个性化效果。
Comments Accepted to ACL 2026