Solving two-armed Bernoulli bandit problems using a Bayesian learning automaton
From MaRDI portal
Recommendations
- On the value of learning for Bernoulli bandits with unknown parameters
- scientific article; zbMATH DE number 4059270
- A learning algorithm for the finite-time two-armed bandit problem
- Optimal Bayesian strategies for the infinite-armed Bernoulli bandit
- Small-sample performance of Bernoulli two-armed bandit Bayesian strategies
Cites work
Cited in
(5)- Thompson sampling guided stochastic searching on the line for deceptive environments with applications to root-finding problems
- Multi-armed bandit for species discovery: a Bayesian nonparametric approach
- Multi-armed bandit problem with online clustering as side information
- Optimal strategy for Bayesian two-armed bandit problem with an arched reward function
- Information-directed sampling for bandits: a primer
This page was built for publication: Solving two-armed Bernoulli bandit problems using a Bayesian learning automaton
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4932958)