Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Correlated Multiarmed Bandit Problem: Bayesian Algorithms and Regret Analysis (1507.01160v2)

Published 5 Jul 2015 in math.OC, cs.LG, and stat.ML

Abstract: We consider the correlated multiarmed bandit (MAB) problem in which the rewards associated with each arm are modeled by a multivariate Gaussian random variable, and we investigate the influence of the assumptions in the Bayesian prior on the performance of the upper credible limit (UCL) algorithm and a new correlated UCL algorithm. We rigorously characterize the influence of accuracy, confidence, and correlation scale in the prior on the decision-making performance of the algorithms. Our results show how priors and correlation structure can be leveraged to improve performance.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (3)
  1. Vaibhav Srivastava (53 papers)
  2. Paul Reverdy (6 papers)
  3. Naomi Ehrich Leonard (61 papers)
Citations (10)

Summary

We haven't generated a summary for this paper yet.