A counterexample on sample-path optimality in stable Markov decision chains with the average reward criterion
From MaRDI portal
(Redirected from Publication:481787)
WRNING: Page is not linked to a MaRDI-Entity. Please add Sitelink / Wikibase-link.