Asymptotically efficient allocation rules for the multiarmed bandit problem with multiple plays-Part II: Markovian rewards (Q3780858): Difference between revisions
From MaRDI portal
Removed claim: author (P16): Item:Q1906866 |
Changed an Item |
||
Property / author | |||
Property / author: Pravin P. Varaiya / rank | |||
Normal rank |
Revision as of 10:46, 1 March 2024
scientific article
Language | Label | Description | Also known as |
---|---|---|---|
English | Asymptotically efficient allocation rules for the multiarmed bandit problem with multiple plays-Part II: Markovian rewards |
scientific article |
Statements
Asymptotically efficient allocation rules for the multiarmed bandit problem with multiple plays-Part II: Markovian rewards (English)
0 references
1987
0 references
learning scheme
0 references
multiarmed bandit
0 references
Markovian rewards
0 references
regret function
0 references