Average cost Markov decision processes with weakly continuous transition probabilities
From MaRDI portal
Abstract: This paper presents sufficient conditions for the existence of stationary optimal policies for average-cost Markov Decision Processes with Borel state and action sets and with weakly continuous transition probabilities. The one-step cost functions may be unbounded, and action sets may be noncompact. The main contributions of this paper are: (i) general sufficient conditions for the existence of stationary discount-optimal and average-cost optimal policies and descriptions of properties of value functions and sets of optimal actions, (ii) a sufficient condition for the average-cost optimality of a stationary policy in the form of optimality inequalities, and (iii) approximations of average-cost optimal actions by discount-optimal actions.
Recommendations
- Average cost Markov decision processes with semi-uniform Feller transition probabilities
- Average cost Markov decision processes: Optimality conditions
- Average optimality for continuous-time Markov decision processes under weak continuity conditions
- Average cost Markov decision processes under the hypothesis of Doeblin
- Partially observable total-cost Markov decision processes with weakly continuous transition probabilities
- Continuous-time Markov decision processes under the risk-sensitive average cost criterion
- Denumerable continuous-time Markov decision processes with multiconstraints on average costs
- Average Reward Markov Decision Processes with Multiple Cost Constraints
- Markov Decision Processes with a Borel Measurable Cost Function—The Average Case
Cites work
- A counterexample on the optimality equation in Markov decision chains with the average cost criterion
- A Counterexample on the Semicontinuity of Minima
- Arbitrary State Markovian Decision Processes
- Average optimality in dynamic programming on Borel spaces -- unbounded costs and controls
- Average Optimality in Dynamic Programming with General State Space
- Compactness of the space of non-randomized policies in countable-state sequential decision processes
- Discrete Dynamic Programming
- Fatou's lemma and Lebesgue's convergence theorem for measures
- Fatou's lemma for weakly converging probabilities
- scientific article; zbMATH DE number 3906790 (Why is no real title available?)
- Markovian Sequential Replacement Processes
- Non-Discounted Denumerable Markovian Decision Models
- On sequential decisions and Markov chains
- On the Nonexistence of $|varepsilon$-Optimal Randomized Stationary Policies in Average Cost Markov Decision Models
- Optimal decision procedures for finite markov chains. Part I: Examples
- Optimality Inequalities for Average Cost Markov Decision Processes and the Stochastic Cash Balance Problem
- OPTIMALITY OF FOUR-THRESHOLD POLICIES IN INVENTORY SYSTEMS WITH CUSTOMER RETURNS AND BORROWING/STORAGE OPTIONS
Cited in
(62)- Average cost Markov decision processes under the hypothesis of Doeblin
- Average cost Markov decision processes: Optimality conditions
- Functional characterization for average cost Markov decision processes with Doeblin's conditions
- On strong average optimality of Markov decision processes with unbounded costs
- Approximation of average cost optimal policies for general Markov decision processes with unbounded costs
- The average cost of Markov chains subject to total variation distance uncertainty
- Solutions of the average cost optimality equation for Markov decision processes with weakly continuous kernel: the fixed-point approach revisited
- Planning for the long run: programming with patient, Pareto responsive preferences
- MDPs with setwise continuous transition probabilities
- On structural properties of optimal average cost functions in Markov decision processes with Borel spaces and universally measurable policies
- Convex analytic method revisited: further optimality results and performance of deterministic policies in average cost stochastic control
- Structure of optimal policies to periodic-review inventory models with convex costs and backorders for all values of discount factors
- Continuity of equilibria for two-person zero-sum games with noncompact action sets and unbounded payoffs
- On the optimality equation for average cost Markov decision processes and its validity for inventory control
- Unbounded dynamic programming via the Q-transform
- On the vanishing discount factor approach for Markov decision processes with weakly continuous transition probabilities
- Berge's maximum theorem for noncompact image sets
- Reduction of total-cost and average-cost MDPs with weakly continuous transition probabilities to discounted mdps
- Another set of conditions for Markov decision processes with average sample-path costs
- A useful technique for piecewise deterministic Markov decision processes
- A survey of average cost problems in deterministic discrete-time control systems
- Partially observable total-cost Markov decision processes with weakly continuous transition probabilities
- Optimality conditions for partially observable Markov decision processes
- New discount and average optimality conditions for continuous-time Markov decision processes
- On the reduction of total-cost and average-cost MDPs to discounted mdps
- A mixed value and policy iteration method for stochastic control with universally measurable policies
- scientific article; zbMATH DE number 4102842 (Why is no real title available?)
- The Existence of a Minimum Pair of State and Policy for Markov Decision Processes under the Hypothesis of Doeblin
- Examples concerning Abel and Cesàro limits
- scientific article; zbMATH DE number 94713 (Why is no real title available?)
- Average Optimality in Dynamic Programming with General State Space
- Convergence of probability measures and Markov decision models with incomplete information
- Continuity of minima: local results
- scientific article; zbMATH DE number 7625164 (Why is no real title available?)
- Stochastic setup-cost inventory model with backorders and quasiconvex cost functions
- Markov decision processes with incomplete information and semiuniform Feller transition probabilities
- Constrained Markov decision processes in Borel spaces: from discounted to average optimality
- Fatou's lemma in its classical form and Lebesgue's convergence theorems for varying measures with applications to Markov decision processes
- Average cost optimality inequality for Markov decision processes with Borel spaces and universally measurable policies
- Average cost Markov decision processes with semi-uniform Feller transition probabilities
- Average optimality for continuous-time Markov decision processes under weak continuity conditions
- Fatou's lemma for weakly converging measures under the uniform integrability condition
- On the Minimum Pair Approach for Average Cost Markov Decision Processes with Countable Discrete Action Spaces and Strictly Unbounded Costs
- Approximation of average cost Markov decision processes using empirical distributions and concentration inequalities
- Uniform Fatou's lemma
- On convergence of value iteration for a class of total cost Markov decision processes
- Formalization of methods for the development of autonomous artificial intelligence systems
- Continuity of discounted values and the structure of optimal policies for <scp>periodic‐review</scp> inventory systems with setup costs
- A note on the existence of optimal stationary policies for average Markov decision processes with countable states
- Another look at partially observed optimal stochastic control: existence, ergodicity, and approximations without belief-reduction
- Epi-consistent approximationof stochastic dynamic programs
- An optimal sequence for sub-Markov decision processes with risk sensitivity
- Acute angle lemma for noncompact image sets
- Near optimal approximations and finite memory policies for POMPDs with continuous spaces
- Discounted cost exponential semi-Markov decision processes with unbounded transition rates: a service rate control problem with impatient customers
- Continuity of filters for discrete-time control problems defined by explicit equations
- AI methodology for modeling protein interactions in biological systems
- The principle of optimality in dynamic programming: a pedagogical note
- Berge's theorem for noncompact image sets
- Multi-objective discounted continuous-time Markov decision processes
- LP based upper and lower bounds for Cesàro and Abel limits of the optimal values in problems of control of stochastic discrete time systems
- Near optimality of quantized policies in stochastic control under weak continuity conditions
This page was built for publication: Average cost Markov decision processes with weakly continuous transition probabilities
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2925348)