Optimal Threshold Policies for Multivariate Stopping-Time POMDPs
From MaRDI portal
Recommendations
- Multiple stopping time POMDPs: structural results \& application in interactive advertising on social media
- Symbolic and Quantitative Approaches to Reasoning with Uncertainty
- Solution Procedures for Partially Observed Markov Decision Processes
- Partially observed Markov decision process multiarmed bandits-structural results
- The problem of optimal stopping in a partially observable Markov chain
Cites work
- Classes of orderings of measures and related correlation inequalities. I. Multivariate totally positive distributions
- Introduction to Stochastic Search and Optimization
- Networked sensor management and data rate control for tracking maneuvering targets
- Partially observed Markov decision process multiarmed bandits-structural results
- Some Monotonicity Results for Partially Observed Markov Decision Processes
- Structural results for partially observed control models
- Structured Threshold Policies for Dynamic Sensor Scheduling—A Partially Observed Markov Decision Process Approach
- Technical Note—On the Convexity of Policy Regions in Partially Observed Systems
Cited in
(5)- Optimal threshold probability in undiscounted Markov decision processes with a target set.
- Multiple stopping time POMDPs: structural results \& application in interactive advertising on social media
- A novel use of value iteration for deriving bounds for threshold and switching curve optimal policies
- Partially observed Markov decision process multiarmed bandits-structural results
- Symbolic and Quantitative Approaches to Reasoning with Uncertainty
This page was built for publication: Optimal Threshold Policies for Multivariate Stopping-Time POMDPs
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q3638204)