Relative value iteration algorithm with soft state aggregation
From MaRDI portal
Recommendations
- Performance Loss Bounds for Approximate Value Iteration with State Aggregation
- Generalized polynomial approximations in Markovian decision processes
- Learning algorithms for Markov decision processes with average cost
- Adaptive aggregation for reinforcement learning in average reward Markov decision processes
- Approximate policy iteration for Markov decision processes via quantitative adaptive aggregations
Cited in
(7)- Reinforcement learning for long-run average cost.
- Revenue management for operations with urgent orders
- Approximate dynamic programming with state aggregation applied to UAV perimeter patrol
- Extreme state aggregation beyond MDPs
- Extreme state aggregation beyond Markov decision processes
- Pseudometrics for State Aggregation in Average Reward Markov Decision Processes
- Performance Loss Bounds for Approximate Value Iteration with State Aggregation
This page was built for publication: Relative value iteration algorithm with soft state aggregation
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2705757)