Asymptotically Efficient Adaptive Choice of Control Laws inControlled Markov Chains
From MaRDI portal
adaptive control of Markov chainscertainty equivalencemultiarmed bandit problemsequential testinguncertainty adjustments
Applications of Markov chains and discrete-time Markov processes on general state spaces (social mobility, learning theory, industrial processes, etc.) (60J20) Sequential statistical analysis (62L10) Adaptive control/observation systems (93C40) Optimal stochastic control (93E20) Stochastic learning and adaptive control (93E35)
Recommendations
- An optimization-oriented approach to the adaptive control of Markov chains
- scientific article; zbMATH DE number 3849083
- Adaptive control of constrained Markov chains: Criteria and policies
- Adaptive control of constrained Markov chains
- scientific article; zbMATH DE number 4042964
- Adaptive control of constrained finite Markov chains
- scientific article; zbMATH DE number 440539
- On adaptive control of a partially observed Markov chain
Cited in
(20)- The multi-armed bandit problem: an efficient nonparametric solution
- Adaptive control design under structured model information limitation: a cost-biased maximum-likelihood approach
- Adaptive policies for perimeter surveillance problems
- Optimal strategies for a class of sequential control problems with precedence relations
- Adaptive control of a Markov chain over a finite parameter set without continuity assumptions on the control laws
- Asymptotically efficient adaptive allocation schemes for controlled Markov chains: finite parameter space
- scientific article; zbMATH DE number 3847250 (Why is no real title available?)
- scientific article; zbMATH DE number 3849083 (Why is no real title available?)
- Assessing the Impact of Head Starts in the Performance of One-Sided Markov-Type Control Schemes
- An asymptotically optimal learning controller for finite Markov chains with unknown transition probabilities
- The Sufficiency of Adjoined Markov Strategies for Controlled Diffusion Processes
- Asymptotically efficient adaptive allocation schemes for controlled i.i.d. processes: finite parameter space
- scientific article; zbMATH DE number 4109753 (Why is no real title available?)
- Minimizing the learning loss in adaptive control of Markov chains under the weak accessibility condition
- scientific article; zbMATH DE number 4123661 (Why is no real title available?)
- Adaptive control of Markov chains with average cost
- Learning the distribution with largest mean: two bandit frameworks
- Learning to optimize via information-directed sampling
- Sequential Generalized Likelihood Ratios and Adaptive Treatment Allocation for Optimal Sequential Selection
- Asymptotically optimal strategies for combinatorial semi-bandits in polynomial time
This page was built for publication: Asymptotically Efficient Adaptive Choice of Control Laws inControlled Markov Chains
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4337732)