arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 1511.08405cs.LGstat.ML

Gains and Losses are Fundamentally Different in Regret Minimization: The Sparse Case

  • Institut de Mathématiques de Jussieu(朱西厄数学研究所)
  • Université Pierre-et-Marie-Curie(皮埃尔和玛丽·居里大学)
  • INRIA(法国国家信息与自动化研究所)
  • Université Paris-Diderot(巴黎第七大学)

机构由 AI 辅助整理,请以论文原文为准。

Joon Kwon, Vianney Perchet

更新

英文摘要:

We demonstrate that, in the classical non-stochastic regret minimization problem with $d$ decisions, gains and losses to be respectively maximized or minimized are fundamentally different. Indeed, by considering the additional sparsity assumption (at each stage, at most $s$ decisions incur a nonzero outcome), we derive optimal regret bounds of different orders. Specifically, with gains, we obtain an optimal regret guarantee after $T$ stages of order $\sqrt{T\log s}$, so the classical dependency in the dimension is replaced by the sparsity size. With losses, we provide matching upper and lower bounds of order $\sqrt{Ts\log(d)/d}$, which is decreasing in $d$. Eventually, we also study the bandit setting, and obtain an upper bound of order $\sqrt{Ts\log (d/s)}$ when outcomes are losses. This bound is proven to be optimal up to the logarithmic factor $\sqrt{\log(d/s)}$.

↑