Reinforcement learning with function approximation: from linear to nonlinear

From MaRDI portal





The subject of this interesting paper is reinforcement learning with function approximation. The subject of reinforcement learning deals with how an agent can learn through interaction with the environment via an optimal policy that maximizes the long-term reward of the agent. When the problem contains a large number of states (often smooth and high-dimensional), a function approximation must be introduced to approximate the involved value or policy functions, and despite their success in many practical applications, for example video, computer vision and many others, a theoretical understanding of reinforcement learning algorithms with function approximation remains relatively limited, particularly when compared to the theoretical results in the tabular setting for the situation of a small number of states, which is much better understood theoretically.\N\NThe paper under review, in particular, reviews recent results on error analysis for reinforcement learning algorithms in linear or nonlinear approximation settings, emphasizing approximation error and estimation error/sample complexity, discussing in particular properties related to approximation error and presenting concrete conditions on reward and transition probability under which these properties hold.\N\NThe paper is well written with a good set of references.



Cites work









This page was built for publication: Reinforcement learning with function approximation: from linear to nonlinear

Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q6955719)