Index Policies and A Novel Performance Space Structure for a Class of Generalised Branching Bandit Problems

Our object of study is a new class of controlled stochastic systems called generalised branching bandits which include discounted branching bandits and generalised bandit problems as special cases. These models allow us to study queueing scheduling and project scheduling problems in which the reward...

Ausführliche Beschreibung

Bibliographische Detailangaben
Veröffentlicht in:Mathematics of Operations Research. - Institute for Operations Research and the Management Sciences. - 25(2000), 2, Seite 281-297
1. Verfasser: Crosbie, J. H. (VerfasserIn)
Weitere Verfasser: Glazebrook, K. D.
Format: Online-Aufsatz
Sprache:English
Veröffentlicht: 2000
Zugriff auf das übergeordnete Werk:Mathematics of Operations Research
Schlagworte:Branching bandit Conservation laws Gittins index Performance space Stochastic scheduling Behavioral sciences Business Applied sciences Mathematics
Beschreibung
Zusammenfassung:Our object of study is a new class of controlled stochastic systems called generalised branching bandits which include discounted branching bandits and generalised bandit problems as special cases. These models allow us to study queueing scheduling and project scheduling problems in which the reward earned from processing a particular job is influenced by other job types present in the system. Our mode of analysis is the achievable region approach. What is novel here is that the full system may be partitioned into two subsystems, each of which satisfies its own set of conservation laws and has its own highly structured performance space. An optimal policy for the full system is constructed from two sets of Gittins indices derived from the conservation laws governing the two subsystems.
ISSN:15265471