A note on generalized second-order value iteration in Markov decision processes
From MaRDI portal
Recommendations
- Acceleration Operators in the Value Iteration Algorithms for Markov Decision Processes
- A First-Order Approach to Accelerated Value Iteration
- Monotone value iteration for discounted finite Markov decision processes
- Value set iteration for Markov decision processes
- An empirical study of policy convergence in Markov decision process value iteration
Cites work
- \({\mathcal Q}\)-learning
- A First-Order Approach to Accelerated Value Iteration
- A Generalized Minimax Q-Learning Algorithm for Two-Player Zero-Sum Stochastic Games
- A Unified Convergence Theory for a Class of Iterative Processes
- Accurately computing the log-sum-exp and softmax functions
- Bias-Corrected Q-Learning With Multistate Extension
- Block monotone iterative methods for numerical solutions of nonlinear elliptic equations
- Convergence Properties of Policy Iteration
- Fixed point theorems in ordered Banach spaces via quasilinearization
- Generalized Second-Order Value Iteration in Markov Decision Processes
- Global Convergence of Newton–Gauss–Seidel Methods
- scientific article; zbMATH DE number 5734462 (Why is no real title available?)
- scientific article; zbMATH DE number 700091 (Why is no real title available?)
- scientific article; zbMATH DE number 2107836 (Why is no real title available?)
- scientific article; zbMATH DE number 3206520 (Why is no real title available?)
- Iterative Solution of Nonlinear Equations in Several Variables
- Monotone Iterations for Nonlinear Equations with Application to Gauss-Seidel Methods
- Newton’s Method for Convex Operators in Partially Ordered Spaces
- On the Convergence of Policy Iteration in Stationary Dynamic Programming
- Reinforcement learning. An introduction
- Solution of a Markovian decision problem by successive overrelaxation
- Value iteration and adaptive dynamic programming for data-driven adaptive optimal control design
This page was built for publication: A note on generalized second-order value iteration in Markov decision processes
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6145054)