Learning-based mean-payoff optimization in an unknown MDP under omega-regular constraints
From MaRDI portal
Recommendations
- Unifying Two Views on Multiple Mean-Payoff Objectives in Markov Decision Processes
- Unifying two views on multiple mean-payoff objectives in Markov decision processes
- PAC Statistical Model Checking of Mean Payoff in Discrete- and Continuous-Time MDP
- Markov decision processes with multiple long-run average objectives
- Energy and Mean-Payoff Parity Markov Decision Processes
Cites work
- scientific article; zbMATH DE number 5869590 (Why is no real title available?)
- scientific article; zbMATH DE number 5685899 (Why is no real title available?)
- scientific article; zbMATH DE number 5585443 (Why is no real title available?)
- Automated technology for verification and analysis. 14th international symposium, ATVA 2016, Chiba, Japan, October 17--20, 2016. Proceedings
- Concurrent games with tail objectives
- Continuity of the value of competitive Markov decision processes
- Deciding parity games in quasipolynomial time
- First-cycle games
- Markov Chains
- Mathematical foundations of computer science 2011. 36th international symposium, MFCS 2011, Warsaw, Poland, August 22--26, 2011. Proceedings
- Meet Your Expectations With Guarantees: Beyond Worst-Case Synthesis in Quantitative Games
- Minimizing expected cost under hard Boolean constraints, with applications to quantitative synthesis
- Multidimensional beyond worst-case and almost-sure problems for mean-payoff objectives
- On the synthesis of strategies in infinite games
- On time with minimal expected cost!
- Permissive strategies: from parity games to safety games
- Pure Stationary Optimal Strategies in Markov Decision Processes
- Reinforcement learning. An introduction
- Robustness of structurally equivalent concurrent parity games
- Shortest paths without a map
- Threshold constraints with guarantees for parity objectives in Markov decision processes
- Tools and algorithms for the construction and analysis of systems. 22nd international conference, TACAS 2016, held as part of the European joint conferences on theory and practice of software, ETAPS 2016, Eindhoven, The Netherlands, April 2--8, 2016. Proc
- Verification of Markov decision processes using learning algorithms
- \({\mathcal Q}\)-learning
Cited in
(9)- Safe learning for near-optimal scheduling
- Omega-Regular Objectives in Model-Free Reinforcement Learning
- Multi-objective -regular reinforcement learning
- PAC Statistical Model Checking of Mean Payoff in Discrete- and Continuous-Time MDP
- Faithful and Effective Reward Schemes for Model-Free Reinforcement Learning of Omega-Regular Objectives
- Model-Free Reinforcement Learning for Lexicographic Omega-Regular Objectives
- scientific article; zbMATH DE number 7559459 (Why is no real title available?)
- scientific article; zbMATH DE number 7559496 (Why is no real title available?)
- PAC statistical model checking of mean payoff in discrete- and continuous-time MDP
This page was built for publication: Learning-based mean-payoff optimization in an unknown MDP under omega-regular constraints
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5009420)