A Note on the Bernoulli Two-Armed Bandit Problem

Thomas A. Kelley

doi:10.1214/aos/1176342827

Password Forgot your password?

Show

Remember Email on this computer

Remember Password

Please wait...

No Project Euclid account? Create an account
or Sign in with your institutional credentials

We can help you reset your password using the email address linked to your Project Euclid account.

Registered users receive a variety of benefits including the ability to customize email alerts, create favorite journals list, and save searches. Please note that a Project Euclid web account does not automatically grant access to full-text content. An institutional or society member subscription is required to view non-Open Access content. Contact customer_support@projecteuclid.org with any questions.
View Project Euclid Privacy Policy

All Fields are Required

* First Name

* Last/Family Name

* Email

* Password

Password Requirements: Minimum 8 characters, must include as least one uppercase, one lowercase letter, and one number or permitted symbol Valid Symbols for password:
~ Tilde
! Exclamation Mark
@ At sign
$ Dollar sign
^ Caret
( Opening Parenthesis
) Closing Parenthesis
_ Underscore
. Period

* Confirm Password

Please wait...

Web Account created successfully

Browse
Resources
About

Advanced Search

Home > Journals > Ann. Statist. > Volume 2 > Issue 5 > Article

September, 1974 A Note on the Bernoulli Two-Armed Bandit Problem

Thomas A. Kelley

Ann. Statist. 2(5): 1056-1062 (September, 1974). DOI: 10.1214/aos/1176342827

ABOUT
FIRST PAGE
CITED BY
DOWNLOAD PAPER SAVE TO MY LIBRARY

PERSONAL SIGN IN
Full access may be available with your subscription

Password Forgot your password?

Show

Remember Email on this computer

Remember Password

No Project Euclid account? Create an account
or Sign in with your institutional credentials

PURCHASE SINGLE ARTICLE

This article is only available to subscribers. It is not available for individual sale.

This will count as one of your downloads.

You will have access to both the presentation and article (if available).

DOWNLOAD NOW

This content is available for download via your institution's subscription. To access this item, please sign in to your personal account.

Password Forgot your password?

Show

Remember Email on this computer

Remember Password

No Project Euclid account? Create an account

My Library

You currently do not have any folders to save your paper to! Create a new folder below.

Abstract

Suppose the arms of a two-armed bandit generate i.i.d. Bernoulli random variables with success probabilities $\rho$ and $\lambda$ respectively. It is desired to maximize the expected sum of $N$ trials where $N$ is fixed. If the prior distribution of $(\rho, \lambda)$ is concentrated at two points $(a, b)$ and $(c, d)$ in the unit square, a characterization of the optimal policy is given. In terms of $a, b, c$, and $d$, necessary and sufficient conditions are given for the optimality of the myopic policy.

Citation

Download Citation

Thomas A. Kelley. "A Note on the Bernoulli Two-Armed Bandit Problem." Ann. Statist. 2 (5) 1056 - 1062, September, 1974. https://doi.org/10.1214/aos/1176342827

Information

Published: September, 1974

First available in Project Euclid: 12 April 2007

zbMATH: 0299.62048

MathSciNet: MR426324

Digital Object Identifier: 10.1214/aos/1176342827

Keywords: 62.35 , 62.45 , Bernoulli random variable , myopic , optimal , posterior distribution , relative advantage , sequential , strategy , two-armed bandit problem , two-point prior distribution

Access the abstract

JOURNAL ARTICLE
7 PAGES

DOWNLOAD PDF + SAVE TO MY LIBRARY