Asymptotically efficient allocation rules for the multiarmed bandit problem with multiple plays-Part II: Markovian rewards (Q3780858): Difference between revisions
From MaRDI portal
Added link to MaRDI item. |
Removed claim: author (P16): Item:Q1906866 |
||
Property / author | |||
Property / author: Pravin P. Varaiya / rank | |||
Revision as of 10:46, 1 March 2024
scientific article
Language | Label | Description | Also known as |
---|---|---|---|
English | Asymptotically efficient allocation rules for the multiarmed bandit problem with multiple plays-Part II: Markovian rewards |
scientific article |
Statements
Asymptotically efficient allocation rules for the multiarmed bandit problem with multiple plays-Part II: Markovian rewards (English)
0 references
1987
0 references
learning scheme
0 references
multiarmed bandit
0 references
Markovian rewards
0 references
regret function
0 references