Approximate Value Iteration with Temporally Extended Actions
From MaRDI portal
Recommendations
- Incremental Value Iteration for Time-Aggregated Markov-Decision Processes
- scientific article; zbMATH DE number 1509479
- On the existence of fixed points for approximate value iteration and temporal-difference learning
- Approximate Value Iteration for Risk-Aware Markov Decision Processes
- Iteratively extending time horizon reinforcement learning.
- Value iteration for long-run average reward in Markov decision processes
- Value and Policy Function Approximations in Infinite-Horizon Optimization Problems
- Complexity bounds for approximately solving discounted MDPs by value iterations
Cited in
(3)
This page was built for publication: Approximate Value Iteration with Temporally Extended Actions
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2941739)