Equilibrium in misspecified Markov decision processes

From MaRDI portal
Publication:5164471

DOI10.3982/TE3843zbMATH Open1475.91181arXiv1502.06901OpenAlexW3157638780MaRDI QIDQ5164471FDOQ5164471


Authors: Ignacio Esponda, Demian Pouzo Edit this on Wikidata


Publication date: 11 November 2021

Published in: Theoretical Economics (Search for Journal in Brave)

Abstract: We study Markov decision problems where the agent does not know the transition probability function mapping current states and actions to future states. The agent has a prior belief over a set of possible transition functions and updates beliefs using Bayes' rule. We allow her to be misspecified in the sense that the true transition probability function is not in the support of her prior. This problem is relevant in many economic settings but is usually not amenable to analysis by the researcher. We make the problem tractable by studying asymptotic behavior. We propose an equilibrium notion and provide conditions under which it characterizes steady state behavior. In the special case where the problem is static, equilibrium coincides with the single-agent version of Berk-Nash equilibrium (Esponda and Pouzo (2016)). We also discuss subtle issues that arise exclusively in dynamic settings due to the possibility of a negative value of experimentation.


Full work available at URL: https://arxiv.org/abs/1502.06901




Recommendations





Cited In (6)





This page was built for publication: Equilibrium in misspecified Markov decision processes

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q5164471)