Structural Properties of Bayesian Bandits with Exponential Family Distributions

Mathematics – Statistics Theory

Scientific paper

Rate now

  [ 0.00 ] – not rated yet Voters 0   Comments 0

Details

Scientific paper

We study a bandit problem where observations from each arm have an exponential family distribution and different arms are assigned independent conjugate priors. At each of n stages, one arm is to be selected based on past observations. The goal is to find a strategy that maximizes the expected discounted sum of the $n$ observations. Two structural results hold in broad generality: (i) for a fixed prior weight, an arm becomes more desirable as its prior mean increases; (ii) for a fixed prior mean, an arm becomes more desirable as its prior weight decreases. These generalize and unify several results in the literature concerning specific problems including Bernoulli and normal bandits. The second result captures an aspect of the exploration-exploitation dilemma in precise terms: given the same immediate payoff, the less one knows about an arm, the more desirable it becomes because there remains more information to be gained when selecting that arm. For Bernoulli and normal bandits we also obtain extensions to nonconjugate priors.

No associations

LandOfFree

Say what you really think

Search LandOfFree.com for scientists and scientific papers. Rate them and share your experience with other people.

Rating

Structural Properties of Bayesian Bandits with Exponential Family Distributions does not yet have a rating. At this time, there are no reviews or comments for this scientific paper.

If you have personal experience with Structural Properties of Bayesian Bandits with Exponential Family Distributions, we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Structural Properties of Bayesian Bandits with Exponential Family Distributions will most certainly appreciate the feedback.

Rate now

     

Profile ID: LFWR-SCP-O-262226

  Search
All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.