Bayesian Inference of Contextual Bandit Policies via Empirical Likelihood
通过经验似然进行上下文老虎机策略的贝叶斯推断
机构 * School of Mathematics and Statistics University of Melbourne(数学与统计学学院墨尔本大学)
AI总结 本文提出了一种基于经验似然的贝叶斯推断方法,用于在有限样本条件下对多个上下文老虎机策略进行联合分析,实现了对策略比较的灵活推断和不确定性量化。
Comments Accepted for publication in JMLR