Some Bounds for Discounted Sequential Decision Processes
From MaRDI portal
(Redirected from Publication:5640967)
Cited in
(32)- Variational characterizations in Markov decision processes
- Nonstationary Markov decision problems with converging parameters
- The method of value oriented successive approximations for the average reward Markov decision process
- Bounds for the renewal function
- Conditions for characterizing the structure of optimal strategies in infinite-horizon dynamic programs
- A natural extension of the MacQueen extrapolation
- Estimates for finite-stage dynamic programs
- Discounted Markov games: Generalized policy iteration method
- Discounted Markov games; successive approximation and stopping times
- Contraction mappings underlying undiscounted Markov decision problems
- A K-step look-ahead analysis of value iteration algorithms for Markov decision processes
- Computational comparison of value iteration algorithms for discounted Markov decision processes
- Block-successive approximation for a discounted Markov decision model
- Serial and parallel value iteration algorithms for discounted Markov decision processes
- Error bounds for stochastic shortest path problems
- Using adaptive learning in credit scoring to estimate take-up probability distribution
- Block-scaling of value-iteration for discounted Markov renewal programming
- Replacement process decomposition for discounted Markov renewal programming
- MARKOV DECISION PROCESSES
- Some basic concepts of numerical treatment of Markov decision models
- (Approximate) iterated successive approximations algorithm for sequential decision processes
- Improved iterative computation of the expected discounted return in Markov and semi-Markov chains
- Computation techniques for large scale undiscounted markov decision processes
- A decision exclusion algorithm for a class of Markovian Decision Processes
- A set of successive approximation methods for discounted Markovian decision problems
- A method of bisection for discounted Markov decision problems
- A superharmonic approach to solving infinite horizon partially observable Markov decision problems
- Zur Extrapolation in Markoffschen Entscheidungsmodellen mit Diskontierung
- An Heuristic for Multi-Dimensional Markov Decision Processes
- Bounds on the fixed point of a monotone contraction operator
- Solving infinite horizon discounted Markov decision process problems for a range of discount factors
- Markov decision processes
This page was built for publication: Some Bounds for Discounted Sequential Decision Processes
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5640967)