Reinforcement learning based algorithms for average cost Markov decision processes (Q2643632): Difference between revisions

From MaRDI portal
Import240304020342 (talk | contribs)
Set profile property.
Set OpenAlex properties.
Property / full work available at URL
 
Property / full work available at URL: https://doi.org/10.1007/s10626-006-0003-y / rank
 
Normal rank
Property / OpenAlex ID
 
Property / OpenAlex ID: W2061769118 / rank
 
Normal rank

Revision as of 00:25, 20 March 2024

scientific article
Language Label Description Also known as
English
Reinforcement learning based algorithms for average cost Markov decision processes
scientific article

    Statements

    Reinforcement learning based algorithms for average cost Markov decision processes (English)
    0 references
    27 August 2007
    0 references
    actor-critic algorithms
    0 references
    two timescale stochastic approximation
    0 references
    Markov decision processes
    0 references
    policy iteration
    0 references
    simultaneous perturbation stochastic approximation
    0 references
    normalized Hadamard matrices
    0 references
    reinforcement learning
    0 references
    TD-learning
    0 references

    Identifiers