An adaptive optimal controller for discrete-time Markov environments
From MaRDI portal
Publication:4152423
Cited in
(6)- The convergence of \(TD(\lambda)\) for general \(\lambda\)
- Risk-averse autonomous systems: a brief history and recent developments from the perspective of optimal control
- Basis function adaptation in temporal difference reinforcement learning
- A Spiking Neural Network Model of an Actor-Critic Learning Agent
- Reinforcement Learning, Bit by Bit
- Mathematical properties of neuronal TD-rules and differential Hebbian learning: a comparison
This page was built for publication: An adaptive optimal controller for discrete-time Markov environments
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q4152423)