A convergent online single time scale actor critic algorithm
From MaRDI portal
Recommendations
Cited in
(6)- Natural actor-critic algorithms
- Natural actor-critic based on batch recursive least-squares
- OnActor-Critic Algorithms
- Global convergence of policy gradient methods to (almost) locally optimal policies
- A Small Gain Analysis of Single Timescale Actor Critic
- On the sample complexity of actor-critic method for reinforcement learning with function approximation
This page was built for publication: A convergent online single time scale actor critic algorithm
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q2896031)