A learning algorithm for the finite-time two-armed bandit problem
From MaRDI portal
Recommendations
Cited in
(10)- A penalized bandit algorithm
- When can the two-armed bandit algorithm be trusted?
- On the value of learning for Bernoulli bandits with unknown parameters
- An Efficient Algorithm for Learning with Semi-bandit Feedback
- How Fast Is the Bandit?
- Machine learning and nonparametric bandit theory
- Solving two-armed Bernoulli bandit problems using a Bayesian learning automaton
- scientific article; zbMATH DE number 7380836 (Why is no real title available?)
- Satisficing in Time-Sensitive Bandit Learning
- On optimal prior learning time in the two-armed bandit problem
This page was built for publication: A learning algorithm for the finite-time two-armed bandit problem
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3342234)