The multi-armed bandit problem under the mean-variance setting
From MaRDI portal
Cites work
- An online algorithm for the risk-aware restless bandit
- Asymptotically efficient adaptive allocation rules
- Exploration-exploitation tradeoff using variance estimates in multi-armed bandits
- Finite-time analysis of the multiarmed bandit problem
- scientific article; zbMATH DE number 3084669 (Why is no real title available?)
- Index policy for multiarmed bandit problem with dynamic risk measures
- Multi-armed bandit-based hyper-heuristics for combinatorial optimization problems
- Multi-Armed Bandits With Correlated Arms
- On the likelihood that one unkrown probability exeeds another in view of the evidence of two samples.
- Sample mean based index policies by O(log n) regret for the multi-armed bandit problem
- Some aspects of the sequential design of experiments
This page was built for publication: The multi-armed bandit problem under the mean-variance setting
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6981676)