Average Cost Dynamic Programming Equations For Controlled Markov Chains With Partial Observations
From MaRDI portal
Recommendations
- Ergodic control of partially observed Markov chains
- Control of Markov Chains with Long-Run Average Cost Criterion: The Dynamic Programming Equations
- scientific article; zbMATH DE number 4085565
- scientific article; zbMATH DE number 4152290
- The value function in ergodic control of diffusion processes with partial observations II
Cited in
(19)- Weak Feller property of non-linear filters
- Successive approximations in partially observable controlled Markov chains with risk-sensitive average criterion
- Long run control with degenerate observation
- A further remark on dynamic programming for partially observed Markov processes
- Strong uniform value in gambling houses and partially observable Markov decision processes
- Control of Markov Chains with Long-Run Average Cost Criterion: The Dynamic Programming Equations
- Isomorphism Properties of Optimality and Equilibrium Solutions Under Equivalent Information Structure Transformations: Stochastic Dynamic Games and Teams
- On the existence of stationary optimal policies for partially observed MDPs under the long-run average cost criterion
- Finite-memory strategies in POMDPs with long-run average objectives
- Zero-sum games involving teams against teams: existence of equilibria, and comparison and regularity in information
- Markov control with rare state observation: average optimality
- Dynamic programming for ergodic control with partial observations.
- Geometry of information structures, strategic measures and associated stochastic control topologies
- Average cost optimality of partially observed MDPs: contraction of nonlinear filters and existence of optimal solutions and approximations
- History-dependent evaluations in partially observable Markov decision process
- Robustness to incorrect priors and controlled filter stability in partially observed stochastic control
- Partially observed semi-Markov zero-sum games with average payoff
- Sequential stochastic control (single or multi-agent) problems nearly admit change of measures with independent measurement
- Another look at partially observed optimal stochastic control: existence, ergodicity, and approximations without belief-reduction
This page was built for publication: Average Cost Dynamic Programming Equations For Controlled Markov Chains With Partial Observations
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4507473)