Multiobjective Bayesian Bandits
P W Jones · 1992
Abstract Modifications of the Bernoulli two armed bandit with finite horizon and independent beta priors are discussed. A terminal decision rule and a stopping rule are introduced. Optimal sequential designs and their characteristics are derived using a series of recurrence relations. The effects of changes in prior information are studied.