On the Speed of Convergence of Value Iteration on Stochastic Shortest-Path Problems
From MaRDI portal
Recommendations
- Error bounds for stochastic shortest path problems
- An Analysis of Stochastic Shortest Path Problems
- On boundedness of Q-learning iterates for stochastic shortest path problems
- The convergence of value iteration in discounted Markov decision processes
- Q-learning and policy iteration algorithms for stochastic shortest path problems
Cited in
(8)- Computing transience bounds of emergency call centers: a hierarchical timed Petri net approach
- The stochastic shortest-path problem for Markov chains with infinite state space with applications to nearest-neighbor lattice chains
- Error bounds for stochastic shortest path problems
- Robust shortest path planning and semicontractive dynamic programming
- Q-learning and policy iteration algorithms for stochastic shortest path problems
- Speed of Convergence and Stopping Rules in an Iterative Planning Procedure for Nonconvex Economies
- On boundedness of Q-learning iterates for stochastic shortest path problems
- Ranking policies in discrete Markov decision processes
This page was built for publication: On the Speed of Convergence of Value Iteration on Stochastic Shortest-Path Problems
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5388035)