On the Optimal Reward Function of the Continuous Time Multiarmed Bandit Problem
From MaRDI portal
Recommendations
Cited in
(16)- Optimal stopping problems for multiarmed bandit processes with arms' independence
- Applicable stochastic control: From theory to practice
- Synchronization and optimality for multi-armed bandit problems in continuous time
- On monotone optimal decision rules and the stay-on-a-winner rule for the two-armed bandit
- Multi-armed bandit processes with optimal selection of the operating times
- scientific article; zbMATH DE number 4064879 (Why is no real title available?)
- scientific article; zbMATH DE number 503440 (Why is no real title available?)
- scientific article; zbMATH DE number 736275 (Why is no real title available?)
- The system of quasi-variational inequalities attached to the two-armed bandit problem
- Regret and Convergence Bounds for a Class of Continuum-Armed Bandit Problems
- A general theory of multiarmed bandit processes with constrained arm switches
- Minimax Off-Policy Evaluation for Multi-Armed Bandits
- Bandit problems with Lévy processes
- Finite-time analysis of the multiarmed bandit problem
- Optimal activation of halting multi‐armed bandit models
- Randomization in the two-armed bandit problem
This page was built for publication: On the Optimal Reward Function of the Continuous Time Multiarmed Bandit Problem
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3200906)