L\'evy bandits under Poissonian decision times

Kazutoshi Yamazaki

L\'evy bandits under Poissonian decision times

Abstract

We consider a version of the continuous-time multi-armed bandit problem where decision opportunities arrive at Poisson arrival times, and study its Gittins index policy. When driven by spectrally one-sided L\'evy processes, the Gittins index can be written explicitly in terms of the scale function, and is shown to converge to that in the classical L\'evy bandit of Kaspi and Mandelbaum (1995).

0

Turn this paper into a lesson

ArcXiv compiles a structured reading guide from this paper's metadata: plain-English importance, contributions, prerequisite concepts, which sections to read first, flashcards, and a quiz. Grounded in the abstract, never invented.

Or open the topic learn hub

Discussion (0)

Sign in to join the discussion.

Loading comments…