Contractivity of Bellman operator in risk averse dynamic programming with infinite horizon
From MaRDI portal
Publication:6161899
Abstract: The paper deals with a risk averse dynamic programming problem with infinite horizon. First, the required assumptions are formulated to have the problem well defined. Then the Bellman equation is derived, which may be also seen as a standalone reinforcement learning problem. The fact that the Bellman operator is contraction is proved, guaranteeing convergence of various solution algorithms used for dynamic programming as well as reinforcement learning problems, which we demonstrate on the value iteration algorithm.
Recommendations
- Risk-averse dynamic programming for Markov decision processes
- Abstract dynamic programming
- Elementary results on solutions to the Bellman equation of dynamic programming: existence, uniqueness, and convergence
- Convex dynamic programming with (bounded) recursive utility
- On discounted dynamic programming with unbounded returns
Cites work
This page was built for publication: Contractivity of Bellman operator in risk averse dynamic programming with infinite horizon
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6161899)