Performance guarantees for empirical Markov decision processes with applications to multiperiod inventory models
From MaRDI portal
Publication:4904590
Recommendations
- On the convergence of optimal actions for Markov decision processes and the optimality of (s,S) inventory policies
- Approximation of average cost Markov decision processes using empirical distributions and concentration inequalities
- Provably Near-Optimal Sampling-Based Policies for Stochastic Inventory Control Models
- Robust Markov Decision Processes with Data-Driven, Distance-Based Ambiguity Sets
- Robust Markov Decision Processes
Cited in
(5)- Empirical dynamic programming
- (s,S) inventory systems with correlated demands
- Some limit properties of Markov chains induced by recursive stochastic algorithms
- Dynamic Inventory Control with Fixed Setup Costs and Unknown Discrete Demand Distribution
- The minimal hitting probability of continuous-time controlled Markov systems with countable states
This page was built for publication: Performance guarantees for empirical Markov decision processes with applications to multiperiod inventory models
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4904590)